A practical starting point
Clean text lists and deduplicate CSV records
Clean repeated list entries or compare CSV records by selected columns. Review kept and removed results, preserve input order, and keep a source copy.
Useful starting points
Tools and resources you can explore today. External services have their own terms and availability.
A practical way forward
Keep the original data
Save an unchanged source copy. Decide whether order matters and whether similar-looking entries could represent different people, items or events.Apply one cleanup at a time
Inspect the blank-line and duplicate counts, remove only the unwanted rows, and review the result before sorting.Verify what remains
Compare the cleaned list with the original. Check accented characters, letter case and spaces because near-matches are different from exact duplicate lines.Use record-aware matching for CSV
A CSV field can contain a comma or newline, so do not treat CSV records as independent text lines. Use the CSV duplicate remover, choose comparison columns and first/last input order, then review both kept and removed records.
Keep in mind
- Exact duplicate removal does not resolve spelling variants or identify the same person.
- Sorting changes order; preserve a copy if sequence has meaning.
- CSV key matching is not fuzzy identity matching. First/last refers to input order, and records with empty selected keys are kept separately by default.
Sources, not guesswork
Background for this guide. Discussion and documentation are useful signals, not proof of demand or an endorsement of every service.
- Duplicate-line handling reference Official guide
Documents duplicate-line semantics; our browser tool has its own behavior.
Planned · Research
Reviewable list cleanup
Research only — not available yet. Would near-duplicate suggestions with an approval step be useful for the lists you handle?
This feature is not available yet. We’re exploring whether it would be useful. Tell us about your task; leaving an email for follow-up is optional.