Wayback Machine Scraper – Internet Archive Snapshots avatar

Wayback Machine Scraper – Internet Archive Snapshots

Pricing

from $1.00 / 1,000 results

Go to Apify Store
Wayback Machine Scraper – Internet Archive Snapshots

Wayback Machine Scraper – Internet Archive Snapshots

$1/1K 🔥 Fast Wayback Machine scraper! Internet Archive snapshot history & availability — timestamps, URLs, status & MIME. JSON, CSV, Excel or API in seconds. Paste URLs & pull thousands of captures for OSINT, SEO & research ⚡

Pricing

from $1.00 / 1,000 results

Rating

0.0

(0)

Developer

ninhothedev

ninhothedev

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

a day ago

Last modified

Share

Get clean snapshot history and closest-available captures for any URL from the Wayback Machine (Internet Archive). Full CDX capture timeline or the snapshot nearest a date. As JSON, CSV or Excel. No key, no proxy.

💵 Pricing

Pay per result — about $1 per 1,000 items (input and output). No subscription, no proxy needed. You only pay for the data you scrape, and new Apify accounts include free monthly credits so you can test it for $0.

🔎 What can it extract?

  • Full snapshot history for a URL (every archived capture via the CDX index)
  • Capture timestamp_iso, ready-to-open snapshot_url, MIME type, HTTP status, digest, length
  • Closest available snapshot to any target date (or the latest)
  • Works with bare domains (apple.com) or deep URLs (github.com/about)
  • Optional fromYear / toYear range filtering

🚀 How do I use it?

  1. Click Try for free.
  2. Set the input (see example below).
  3. Click Start and download results as JSON, CSV or Excel — or pull them via the API / schedule them.

⚙️ Input example

{
"mode": "snapshots",
"urls": ["apple.com", "github.com"],
"maxItems": 100
}

Availability lookup:

{
"mode": "available",
"urls": ["apple.com"],
"targetDate": "20150101"
}

📦 Output example

{
"url": "http://www.apple.com/",
"timestamp_iso": "1998-05-25T09:12:35",
"snapshot_url": "http://web.archive.org/web/19980525091235/http://www.apple.com/",
"mimetype": "text/html",
"status_code": "200"
}

🎯 Use cases

  • OSINT — investigate how a site looked and changed over time
  • SEO — audit historical content, redirects and status codes
  • Research — build datasets of web history and page evolution
  • Web archiving — enumerate every capture of a domain

💰 How much will it cost?

You pay only the per-result price (no proxy cost): 100 → ~$0.10 · 1,000 → ~$1 · 10,000 → ~$10

🆚 Why this one

FeatureThis actorOthers
Full CDX capture historylatest only
Closest-to-date lookuppartial
Bare domain + deep URLsometimes
No key, no proxysometimes

❓ FAQ

Key/proxy? No, the Internet Archive is public. How far back? As far as the archive holds — often the late 1990s. Bulk? Pass many URLs at once.

🛟 Support

Found a bug or need an extra field? Open an issue on the actor — fixes and new fields ship fast.

Data via the public Internet Archive Wayback Machine (availability + CDX APIs).


Keywords: wayback machine scraper, internet archive, web archive, snapshot history, cdx, osint, seo, domain history, JSON CSV Excel