Wayback Machine Snapshot History Scraper
Pricing
Pay per event
Wayback Machine Snapshot History Scraper
Export Internet Archive Wayback Machine snapshot history with replay URLs, timestamps, status, MIME, digest, and size filters.
Wayback Machine Snapshot History Scraper
Pricing
Pay per event
Export Internet Archive Wayback Machine snapshot history with replay URLs, timestamps, status, MIME, digest, and size filters.
One to 100 targets. Examples: nasa.gov, https://www.nasa.gov/history/, or nasa.gov/*.
[ "nasa.gov"]Domain includes subdomains; host stays on one host; prefix finds URLs below a prefix; exact matches one URL.
Optional UTC boundary using 4–14 digits, from YYYY to YYYYMMDDhhmmss.
Optional inclusive UTC boundary using 4–14 digits, from YYYY to YYYYMMDDhhmmss.
Exact status codes (200, 301) or classes (2.., 3..). All selected filters must match.
[ "200"]CDX MIME types such as text/html, application/pdf, or image/*.
[ "text/html"]Return only one capture for repeated records with the same Internet Archive digest.
Maximum dataset records across all targets.
Records requested per CDX page. Lower this if the archive rate-limits large queries.