How to use the Remove Duplicate Lines
- 1
Paste your list into the Input pane, one item per line, or press Sample. The deduplicated list appears on the right as you type.
- 2
Choose what to do with repeats. Remove duplicates keeps one copy of each line, Show only duplicates lists the values that repeat, and Count occurrences shows how often each line appears.
- 3
Tick Ignore case to treat Apple and apple as the same, and Ignore surrounding spaces to treat " apple" and "apple" as the same. Choose whether to keep the first or the last copy.
- 4
Open the Sort or Clean up panel to sort the result, drop empty lines, trim every line, or add a prefix and suffix such as quotes and commas.
- 5
Copy the result, download it, or press the arrow to move it back into the input and keep working.
Features
- Removes duplicate lines instantly while keeping the original order
- Case-insensitive and whitespace-insensitive comparison
- Keep the first or the last occurrence of each value
- Show only the lines that repeat, or count how often each line occurs
- Optional sorting, trimming, empty-line removal, numbering, prefix and suffix in the same pass
- Status line with lines in, lines out and duplicates removed
- Handles hundreds of thousands of lines in your browser
- Nothing is uploaded; your lists stay on your device
Clean lists without a spreadsheet
Duplicate lines creep into every list that is assembled from more than one place: email addresses exported from two systems, URLs collected from logs, keywords merged from several research tools, product codes copied from invoices. Spreadsheets can remove them, but only after importing, choosing a column and confirming a dialog. Here you paste the list and the unique version is already on the right.
The comparison is exact by default and can be relaxed in two ways. Ignore case folds upper and lower case together, which is what you want for email addresses and tags. Ignore surrounding spaces ignores the padding that copy and paste tends to add. Both affect only the comparison: the copy that is kept is written exactly as it appeared.
Finding repeats instead of removing them
Sometimes the duplicates are the interesting part. Show only duplicates lists every value that occurs more than once, which is the quickest way to find a double booking, a reused ID or a customer who signed up twice. Count occurrences turns the list into a frequency table, with the most common lines first, ready to paste into a spreadsheet as two columns.
Related tasks
To put the cleaned list in order, use Sort Lines, which shares the same pipeline with sorting opened first. To change many values at once rather than remove them, use Find and Replace.
Frequently asked questions
Does removing duplicates change the order of my list?
Why are some lines that look the same not removed?
How do I find which lines are duplicated?
Can I remove empty lines at the same time?
How large a list can I deduplicate?
Is my data uploaded?
Last updated .