Wayback Machine Snapshot & Page Change Tracker
Pricing
from $1.80 / 1,000 results
Wayback Machine Snapshot & Page Change Tracker
Wayback Machine Snapshot & Page Change Tracker lists every Internet Archive capture of a URL and diffs the visible text of the oldest vs. newest snapshot in range — one scored change row per URL, or a full snapshot list.
Pricing
from $1.80 / 1,000 results
Rating
0.0
(0)
Developer
Murat Uzun
Maintained by CommunityActor stats
0
Bookmarked
3
Total users
2
Monthly active users
2 days ago
Last modified
Categories
Share
What is Wayback Machine Snapshot & Page Change Tracker?
Wayback Machine Snapshot & Page Change Tracker is an Apify Actor that queries the Internet Archive's Wayback Machine for every capture of a URL and turns that history into structured data. In diff mode (default) it fetches the oldest and newest archived snapshot in your date range and returns one row per URL describing exactly what changed: title, word count, a line-based text diff with sample added/removed lines, and link churn. In snapshots mode it simply lists every capture, newest first, so you can pick specific dates yourself. No API key, no proxy, no headless browser — just the Wayback Machine's own CDX index and its id_ raw-snapshot endpoint.
Why use Wayback Machine Snapshot & Page Change Tracker?
- Pricing and terms history — prove what a vendor's pricing page or ToS said on a given date, for legal, compliance or renewal-negotiation purposes.
- SEO history — see what a page's title, headings and copy looked like before a ranking change or a competitor's redesign.
- Competitor and brand monitoring — track when a competitor rewrote their homepage or landing page, without visiting the site yourself.
- Content audits — spot-check how much a page has actually changed since it was last reviewed, with a percentage score instead of a manual read-through.
How to use Wayback Machine Snapshot & Page Change Tracker
- Paste the page URLs you want history for into URLs, e.g.
https://apify.com/pricing. - Leave Mode on
diffto compare the oldest and newest capture, or switch tosnapshotsto list every capture instead. - Optionally set From date / To date (e.g.
2023-01-01) to restrict the range considered. - Click Start, then export the results as JSON, CSV, Excel or HTML from the Output tab.
Example input
{"urls": ["https://apify.com/pricing", "https://example.com"],"mode": "diff","maxConcurrency": 3}
Example output
Diff mode (one row per URL):
{"url": "https://apify.com/pricing","snapshotCount": 165,"firstSnapshotAt": "2017-10-27T07:20:08.000Z","lastSnapshotAt": "2026-09-07T22:14:50.000Z","oldSnapshotUrl": "https://web.archive.org/web/20171027072008/https://www.apify.com/pricing","newSnapshotUrl": "https://web.archive.org/web/20260907221450/https://apify.com/pricing","oldTitle": "Pricing","newTitle": "Apify pricing - flexible plan + pay as you go · Apify","titleChanged": true,"oldWordCount": 570,"newWordCount": 2084,"changedPercent": 95.3,"addedLines": ["Apify pricing - flexible plan + pay as you go · Apify", "Skip to content", "…"],"removedLines": ["Pricing", "Flexible pricing. Free for developers,", "…"],"addedLinkCount": 187,"removedLinkCount": 9,"error": null,"scrapedAt": "2026-09-12T17:30:00.000Z"}
Snapshots mode (one row per capture):
{"url": "https://example.com/","timestamp": "20260912162353","snapshotAt": "2026-09-12T16:23:53.000Z","snapshotUrl": "https://web.archive.org/web/20260912162353/https://example.com/","digest": "PKUMGV5XIIUJG5CKD4HZMCRWKUMVP5S6","lengthBytes": 1048,"error": null,"scrapedAt": "2026-09-12T17:30:00.000Z"}
You can download the dataset in various formats such as JSON, HTML, CSV, or Excel.
Data table
| Field | Type | Description |
|---|---|---|
url | string | The input URL checked |
snapshotCount, firstSnapshotAt, lastSnapshotAt | number, date | Diff mode: how many captures exist in range and when the first/last were made |
oldSnapshotUrl, newSnapshotUrl | link | Wayback pages for the oldest and newest capture compared |
oldTitle, newTitle, titleChanged | string, boolean | <title> of each capture and whether it differs |
oldWordCount, newWordCount | number | Visible-text word counts |
changedPercent | number | 0-100 line-based change score between the two captures |
addedLines, removedLines | array | Up to 50 sample lines added / removed |
addedLinkCount, removedLinkCount | number | Links present only in the new / only in the old capture |
timestamp, snapshotAt, snapshotUrl, digest, lengthBytes | - | Snapshots mode: one capture's Wayback timestamp, ISO date, URL, content hash and size |
error, scrapedAt | string, date | Per-URL failure reason (never throws the whole run) and when the row was produced |
Input parameters
| Parameter | Type | Default | Description |
|---|---|---|---|
urls | array | ["https://apify.com/pricing"] | Pages to look up, one result group per URL |
mode | string | diff | diff compares oldest vs. newest capture; snapshots lists them all |
from | string | (none) | Only captures on/after this date, e.g. 2023-01-01 |
to | string | (none) | Only captures on/before this date, e.g. 2024-06-30 |
maxSnapshots | integer | 50 | Snapshots mode: max captures returned per URL (1-1000) |
maxConcurrency | integer | 3 | URLs processed in parallel (the run also self-throttles to ~1 req/s) |
Pricing
Wayback Machine Snapshot & Page Change Tracker uses pay-per-event pricing: $0.003 per result row, i.e. $3 per 1,000 rows, plus a negligible actor-start fee. A diff-mode URL costs one row (two HTML fetches under the hood); snapshots mode charges one row per capture returned, capped by Max snapshots per URL. Set Maximum cost per run and the Actor trims the URL list to what the budget covers.
Wayback Machine Snapshot & Page Change Tracker vs. manually browsing web.archive.org
Clicking through web.archive.org's calendar UI one date at a time does not scale past a handful of pages, and comparing two captures by eye misses small copy changes. This Actor queries the same public CDX index Wayback's own UI uses, but returns a flat, scored dataset for as many URLs as you give it — with a diff percentage, sample changed lines and link churn ready for a spreadsheet, an alert, or an LLM prompt.
Using Wayback Machine Snapshot & Page Change Tracker with AI agents and MCP
This Actor is pay-per-event with limited permissions — the two requirements for an Actor to be callable through the Apify MCP server at mcp.apify.com. An agent passes urls and gets back a structured before/after diff it can summarize or act on, without ever touching a browser or writing scraping code. The same run works from n8n, Make, Zapier and LangChain through Apify's integrations.
FAQ
Why did I get error: "Only one snapshot in range"? The Wayback Machine has archived that URL only once inside your from/to window, so there is nothing to compare it against — widen the date range or drop it.
Why did I get a 429 or 503 error? The Wayback Machine rate-limits aggressively under load. This Actor already self-throttles to about one request per second and retries twice; lowering Max concurrency to 1-2 or simply re-running later usually clears it.
Does this see JavaScript-rendered content? No — it diffs the raw HTML bytes the Wayback Machine stored (via the id_ identity endpoint, no toolbar injection), the same as what the crawler that made the capture saved. Content injected client-side after page load was never archived and cannot be recovered.
Is this legal to run? Yes. The Internet Archive publishes these captures as a public historical record through an open API; no login-gated or paywalled content is accessed.
What are the limitations? Only pages the Wayback Machine actually captured are available — obscure or robots.txt-blocked URLs may have few or no captures. changedPercent is a plain line-based text diff, not a semantic one, so a full re-layout with identical copy can still score high.
Related Actors
Part of the webdatatools web-intelligence suite — every Actor is pay-per-event, reads public data without a login, and returns one clean row per entity:
Browse the whole suite at webdatatools, or call ten of these Actors straight from Claude, Cursor or Cline with the webdatatools MCP server.
Website & domain intelligence
- Email Extractor — Website Contact & Social Finder — e-mails, phones and social profiles per domain
- Tech Stack Detector — Wappalyzer & BuiltWith Alternative — CMS, e-commerce, analytics, pixels and payments per domain
- Domain DNS & Email Security Checker — SPF, DKIM, DMARC, MX provider, registrar and domain age
- Domain Security Audit (TLS, HTTP headers, redirects, robots) — TLS expiry, security headers, redirect chain, robots and llms.txt
- Subdomain Finder (Certificate Transparency) — every subdomain seen in CT logs, with a live DNS check
- Bulk Core Web Vitals & PageSpeed Audit — Lighthouse scores, LCP, CLS, INP and top fixes per URL
- On-Page SEO Audit — title, meta, headings, links, images and schema issues per page
- Sitemap URL Extractor & Change Monitor — every sitemap URL, or new and removed pages between runs
- Bulk Domain WHOIS & RDAP Lookup — registrar, dates, status and nameservers per domain
- Web Scraper — CSS Selector & Data Extractor — pull any CSS selector off any page, one row per URL
- Website Screenshot Generator — full-page or viewport PNG/JPEG screenshots of any URL
Content for AI, LLMs and RAG
- AI Web Search & Read: Google results as clean Markdown — a query turned into clean Markdown from the top search results
- Website to Markdown — Content Crawler for LLM & RAG — any site as clean Markdown per page, no browser
- Article & News Extractor (clean text, author, date, markdown) — clean article text, author, date and Markdown per URL
- Structured Data & JSON-LD Extractor (Schema.org, Open Graph) — Schema.org and Open Graph data from any page
- Google News Scraper (RSS search by keyword, topic, site) — news results by keyword, topic or site
- Press Release Monitor: PR Newswire, BusinessWire, GlobeNewswire — PR Newswire, Business Wire and GlobeNewswire releases
Search, video and social
- YouTube Shorts Scraper — Shorts from channels, hashtags and searches with view counts
- Pinterest Pins Scraper — latest pins of public Pinterest profiles and boards
- YouTube Transcript Scraper — captions and subtitles as text + timed segments, per video or channel
- Google Search Results Scraper — SERP API — organic SERP results per keyword and country
- YouTube Comments Scraper — Comments & Replies — comments and replies with likes, no API key
- YouTube Channel Latest Videos (RSS, no API key) — the latest 15 videos of any channel from RSS
- YouTube Channel Scraper (videos, shorts, live) — a channel's full video, shorts and stream list
- YouTube Search Results Scraper (videos, channels, no API key) — videos, channels and playlists per query
- YouTube Video Details Scraper (views, likes, description, tags) — views, likes, description, tags and chapters per video
- Apple Podcasts Lookup & Episodes Scraper — podcast metadata and episodes from iTunes and RSS
- Bluesky Post, Search & Profile Scraper — posts, profiles, followers and threads from the AT Protocol API
- Telegram Channel Posts Scraper — posts, views and media flags from any public channel
- Substack Publication & Posts Scraper — archive, authors and paywall status per publication
- Google Play Reviews Scraper — reviews, ratings, replies and app versions per app
- App Store Reviews Scraper — iOS reviews and ratings per app and country
- Google Trends Scraper — interest over time, by region, and related queries per keyword
- Google Ads Transparency Scraper — ads any advertiser runs on Google, with format and dates
- Keyword Suggestions Scraper (Google, YouTube, Amazon, Bing) — autocomplete keyword ideas from four search engines
- Bilibili Scraper (Videos, Search, Popular) — Chinese video platform: views, likes, coins, danmaku, uploader
- Mastodon Scraper (Hashtags, Accounts, Trending) — public fediverse posts by hashtag, account or trending
- Meetup Events Scraper (Search by Keyword & City) — upcoming events with RSVPs, fees, venues and groups
- Eventbrite Scraper (Events by Keyword & City) — events by keyword and city with venue, dates and organizer
Leads, jobs and company data
- Career Site Jobs API (Greenhouse, Lever, Ashby, Workday +1) — company domains in, their open jobs out, ATS detected automatically
- Workday Jobs Scraper — jobs with full descriptions from any Workday career site
- Google Maps Scraper — businesses with phone, website, address, rating and coordinates per search
- LinkedIn Jobs Scraper — job titles, companies, locations and full descriptions from LinkedIn job search
- Company 360: full company profile from a domain — one row per domain: contacts, tech, security, hiring and company facts
- Hiring Signals Scraper (Greenhouse, Lever, Ashby, Workable) — open jobs and hiring velocity from 10 public ATS boards
- Y Combinator Companies & Founders Scraper — YC startups by batch, industry and hiring status
- Wikidata Entity & Company Enrichment (facts, IDs, links) — HQ, founders, employees, revenue and social IDs per company
- Email Validator & Verifier — Bulk Email Check — syntax, MX, disposable, role and free-provider checks
- OpenStreetMap POI Extractor (Overpass API: shops, amenities) — shops and amenities by radius, bbox or area
- Stock, Crypto & FX Quotes — one row per symbol from Yahoo, Binance and ECB rates
- Remote Jobs Aggregator (RemoteOK, WWR, Hacker News) — one clean row per remote job, de-duplicated across feeds
- Greenhouse Jobs Scraper — jobs with descriptions from any Greenhouse job board
- Lever Jobs Scraper — jobs with descriptions from any Lever careers page
- Ashby Jobs Scraper — jobs, salaries and descriptions from any Ashby job board
- SmartRecruiters Jobs Scraper — jobs with descriptions from any SmartRecruiters company
- Seek Jobs Scraper (Australia & New Zealand) — Seek job ads with salary, work type and location
- Dice Jobs Scraper — US tech jobs from Dice with salary and remote flag
- AutoScout24 Scraper — European car listings with price, mileage and seller
- Rightmove Scraper — UK property for sale or rent with price and agent
- Wellfound Jobs Scraper (AngelList Startup Jobs) — startup jobs with salary and equity ranges, company size and stage
- Yandex Maps Scraper (Places, Ratings, Phones) — businesses in Russia, Türkiye and the CIS with phones, ratings, hours
- Craigslist Scraper (Listings, Prices, Locations) — listings in any area and category with price, date and coordinates
- JobStreet Scraper (Malaysia, Singapore, PH, ID + JobsDB) — JobStreet and JobsDB jobs in 6 Asian countries with parsed salaries
- InfoJobs Scraper (Spain Jobs, Salaries, Companies) — Spanish jobs with salary range, contract type and full description
- Redfin Scraper (Homes for Sale, Prices, Details) — US homes for sale or sold from any Redfin search, with price and details
- Kleinanzeigen Scraper (Ads, Prices, Locations) — German classifieds with price, VB flag, ZIP, city and seller type
Developer, app and research data
- npm, PyPI & Crates.io Package Health Checker — releases, downloads, deprecation and a health score
- GitHub Repository Health & Activity Report — stars, commits, contributors and risk flags per repo
- VS Code Marketplace Extension Scraper (installs, ratings) — installs, ratings and versions per extension
- Chrome Web Store Extension Scraper (installs, ratings) — users, rating, version and developer per extension
- Google Play Scraper — apps, ratings, installs, developer contact and reviews
- App Store (iOS) App Metadata, Ratings & Top Charts Lookup — ratings, price, version and charts per app
- CrossRef DOI & Citation Metadata Lookup — papers, authors, journals and citation counts
- FDA Recalls & Adverse Events Monitor (openFDA) — food, drug and device recalls from openFDA
- iCal / ICS Calendar Feed to Events Extractor — any public calendar feed as event rows
- Shopify Store Products Scraper — catalog, prices, variants and stock per store
- Hacker News Search & Front Page Scraper — stories, comments and points by query or front page
- GitHub Trending Repositories Scraper — trending repos and developers by language and period
- Stack Overflow & Stack Exchange Q&A Scraper — questions, answers and scores by query, tag or site
- Bulk Image Downloader — download image URLs to storage with size, dimensions and a ZIP
- Google Flights Scraper (Prices, Airlines, Stops) — flight prices, airlines, times, stops and CO2 by route and date
- Google Hotels Scraper (Prices, Ratings, Reviews) — hotel prices per night, stars, rating and reviews by city and dates
- AliExpress Scraper (Search Products & Prices) — AliExpress search results with USD price, discount and rank
- Lazada Scraper (Products, Prices, Sold, Ratings) — Lazada products in 6 countries with price, rating, units sold and seller
Support and feedback
Found a page the Wayback Machine should have captured but this Actor missed, or a diff that looks wrong? Open an issue on the Issues tab.