CCleanMySheet

CleanMySheet guide

Find near-duplicate rows (fuzzy matching)

CleanMySheet groups similar rows using normalized Levenshtein similarity (default ~85%) after blocking on a short key prefix so large files stay usable. Matches appear in a review drawer; nothing is removed until you mark a pair as a duplicate.

Open the cleaner

Steps

  1. Run a scan after upload
  2. Add Fuzzy duplicate detection from search or Recommended
  3. Open Review fuzzy pairs
  4. Mark Duplicate or Keep both
  5. Undo anytime

Not an LLM guess

Similarity is a string distance, not a model hallucination. You can undo the whole step from the recipe list.

Related searches this page answers

  • fuzzy duplicate finder
  • near duplicate rows
  • similar contacts csv