Go to example tasks
Every archived page under one section of a site
Created by
丂卩ㄖㄖҜㄚ
Set matchType to prefix and the run returns every URL the archive holds beneath one path, with capture dates and status for each.
Wayback Machine Scraperspookyweb/wayback-machine-scraper
Captured
Original URL
HTTP
Type
+3 fieldsTextNumberBooleanListObject
Input
URL:https://www.bbc.co.uk/news/
Earliest capture:2022
Latest capture:2026
What to match:prefix
How many captures to keep:none
Maximum snapshots per URL:2000
Detect changes between captures:false
Retrieve the archived page:false
Output fields
Captured
Original URL
HTTP
Type
Content
Text chars
Snapshot
Sign up on Apify01
Create your Apify account to access the Wayback Machine Scraper.
Start the run02
The Actor will start running based on the input automatically.
Receive the output03
Monitor the progress in real-time. You will be notified as soon as your dataset is complete and ready for review.
Integrate into your workflow04
The final output is delivered in JSON, CSV, or Excel format, ready to be plugged into your workflow.
