Deduplicator


Deduplicator — Free Download. Duplicate remover

Simply paste your list into Deduplicator and press the Find Duplicates button. Deduplicator scans the input, detects repeated entries, and displays all matches in a separate preview window. After reviewing the results, press Remove Duplicates to complete the operation. The output list then contains only unique elements. A fuzziness slider gives control over how strictly entries are compared, ranging from exact matching to case-insensitive, spelling-insensitive, and greatly relaxed comparison modes, so operators can tune the deduplication level to suit the data at hand.

★★★★★
4.0(1 ratings)
File size: 12.9 MB
The latest version of Deduplicator is: 2021.06.10
Operating system: Mac OS
Languages: English
Price: $0.00 USD (Free product)
  • Find Duplicates. This function scans the entire pasted list and identifies every entry that appears more than once. It compares items according to the current fuzziness setting, builds a set of repeated values, and presents them in a preview window. The operator can inspect which items were flagged before any changes are applied to the working list.
  • Remove Duplicates. This function deletes all entries that were previously identified as duplicates. It leaves a single instance of each repeated value and preserves the original order of remaining items. The action is performed on the whole list at once, producing a clean output that contains only unique elements.
  • Duplicate Preview Window. This function opens a separate panel that lists every detected duplicate. It shows the repeated values so the operator can verify that the matching rules produced correct results. The preview acts as a checkpoint between detection and removal, reducing the chance of unwanted deletions.
  • Fuzziness Slider. This function adjusts the strictness of the comparison algorithm. At the strict position, entries must match exactly. One step toward fuzzy makes the comparison case-insensitive, so DAVE and Dave are treated as the same value. A further step makes it spelling-insensitive, so Jenni and Jenny are treated as duplicates. The far fuzzy position greatly relaxes matching, so Bob and Bobby are considered equal.
  • Case-Insensitive Matching. This function compares entries without regard to uppercase and lowercase letters. It is activated by moving the fuzziness slider one notch away from strict. The mode is suited to lists where capitalisation varies inconsistently but the underlying values are the same.
  • Spelling-Insensitive Matching. This function treats entries as duplicates when they differ by small spelling variations. It is activated by moving the fuzziness slider another notch away from strict. The mode handles minor typographical differences such as Jenni and Jenny, which would otherwise remain as separate items.
  • Relaxed Matching Mode. This function applies the least strict comparison available in the slider range. It is activated by moving the fuzziness slider all the way toward fuzzy. The mode greatly reduces matching strictness, so shortened or extended forms such as Bob and Bobby are treated as duplicates. Operators can experiment across the slider range to find a satisfactory level.
  • List Input Area. This function provides the field where the source list is pasted. It accepts plain text with one entry per line. All subsequent operations, including detection, preview, and removal, work on the content of this area.
  • Unique Output List. This function presents the deduplicated result after duplicates have been removed. It contains one instance of every value from the original input, with repeated entries eliminated. The output can be copied for use in other documents or applications.
  • Order Preservation. This function keeps the original sequence of the surviving entries. Items are not sorted or rearranged during deduplication. The first occurrence of each value is retained, and later repetitions are discarded, so the resulting list follows the same order as the source.
  • Comparison Range Control. This function defines the span between strict and fuzzy behaviour. Moving the slider across this range changes how aggressively entries are matched. The operator selects a level that fits the nature of the list, balancing between exact identity and approximate similarity.

Deduplicator was created by David Gradwell and is distributed through his software site. The program is written in Java, which allows it to run on multiple operating systems through a Java runtime. The title has been maintained and updated over the years as part of a broader collection of small utilities released by the same developer, alongside tools such as Folder Snapshot Utility, Truck, and Rsync Server. Version notes for those companion programs appear on the same site, documenting releases from 2019 through 2021, including rewritten components and added options.

Alternatives to Deduplicator: