Go to example tasks
Deduplicate site images by content hash for agent memory
Created by
Dennis
Crawl a site and return only unique images (SHA-256 content-hash dedup) — useful for agent memory stores that must avoid duplicate visual entries.
Bulk Image Scraper — Real Dimension & File Size Filtercodeclouds/bulk-image-scraper
Image URL
Found on page
Found via
URL extension
+5 fieldsTextNumberBooleanListObject
Input
Start URLs(required):https://example.com
Max crawl depth:3
Max total images:200
Download full image files:false
Output fields
Image URL
Found on page
Found via
URL extension
Real format
Width (px)
Height (px)
File size (bytes)
Alt text
Sign up on Apify01
Create your Apify account to access the Bulk Image Scraper — Real Dimension & File Size Filter.
Start the run02
The Actor will start running based on the input automatically.
Receive the output03
Monitor the progress in real-time. You will be notified as soon as your dataset is complete and ready for review.
Integrate into your workflow04
The final output is delivered in JSON, CSV, or Excel format, ready to be plugged into your workflow.
