Dataset Deduplicator & Cleaner
Pricing
Pay per usage
Dataset Deduplicator & Cleaner
Remove duplicate rows from an Apify Dataset or inline JSON. Deduplicate by selected fields, trim strings, and write the cleaned rows to a new Dataset.
Dataset Deduplicator & Cleaner
Pricing
Pay per usage
Remove duplicate rows from an Apify Dataset or inline JSON. Deduplicate by selected fields, trim strings, and write the cleaned rows to a new Dataset.
Optional Apify Dataset to clean. If empty, the inline sample rows below are used.
JSON rows to clean when no Source Dataset is selected. The defaults intentionally contain one duplicate for a quick smoke test.
[ { "id": "A001", "name": " Alpha ", "score": 10 }, { "id": "A002", "name": "Beta", "score": 20 }, { "id": "A001", "name": "Alpha duplicate", "score": 99 }]Rows with the same values in these fields are treated as duplicates. Leave empty to compare the whole row.
[ "id"]Choose whether the first or last row wins when duplicate keys are found.
Remove leading and trailing whitespace from top-level string fields before deduplication and output.
When enabled, text values used in deduplication keys are compared without letter-case differences.