Wayback Machine Scraper – Internet Archive Snapshots
Pricing
from $1.00 / 1,000 results
Wayback Machine Scraper – Internet Archive Snapshots
$1/1K 🔥 Fast Wayback Machine scraper! Internet Archive snapshot history & availability — timestamps, URLs, status & MIME. JSON, CSV, Excel or API in seconds. Paste URLs & pull thousands of captures for OSINT, SEO & research ⚡
Pricing
from $1.00 / 1,000 results
Rating
0.0
(0)
Developer
ninhothedev
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
a day ago
Last modified
Categories
Share
Get clean snapshot history and closest-available captures for any URL from the Wayback Machine (Internet Archive). Full CDX capture timeline or the snapshot nearest a date. As JSON, CSV or Excel. No key, no proxy.
💵 Pricing
Pay per result — about $1 per 1,000 items (input and output). No subscription, no proxy needed. You only pay for the data you scrape, and new Apify accounts include free monthly credits so you can test it for $0.
🔎 What can it extract?
- Full snapshot history for a URL (every archived capture via the CDX index)
- Capture
timestamp_iso, ready-to-opensnapshot_url, MIME type, HTTP status, digest, length - Closest available snapshot to any target date (or the latest)
- Works with bare domains (
apple.com) or deep URLs (github.com/about) - Optional
fromYear/toYearrange filtering
🚀 How do I use it?
- Click Try for free.
- Set the input (see example below).
- Click Start and download results as JSON, CSV or Excel — or pull them via the API / schedule them.
⚙️ Input example
{"mode": "snapshots","urls": ["apple.com", "github.com"],"maxItems": 100}
Availability lookup:
{"mode": "available","urls": ["apple.com"],"targetDate": "20150101"}
📦 Output example
{"url": "http://www.apple.com/","timestamp_iso": "1998-05-25T09:12:35","snapshot_url": "http://web.archive.org/web/19980525091235/http://www.apple.com/","mimetype": "text/html","status_code": "200"}
🎯 Use cases
- OSINT — investigate how a site looked and changed over time
- SEO — audit historical content, redirects and status codes
- Research — build datasets of web history and page evolution
- Web archiving — enumerate every capture of a domain
💰 How much will it cost?
You pay only the per-result price (no proxy cost): 100 → ~$0.10 · 1,000 → ~$1 · 10,000 → ~$10
🆚 Why this one
| Feature | This actor | Others |
|---|---|---|
| Full CDX capture history | ✅ | latest only |
| Closest-to-date lookup | ✅ | partial |
| Bare domain + deep URL | ✅ | sometimes |
| No key, no proxy | ✅ | sometimes |
🔗 Related actors
- Website Tech Stack Detector — detect frameworks & tools
- RDAP Domain Scraper — domain registration data
- Google News Scraper — news articles
- Wikipedia Scraper — knowledge
❓ FAQ
Key/proxy? No, the Internet Archive is public. How far back? As far as the archive holds — often the late 1990s. Bulk? Pass many URLs at once.
🛟 Support
Found a bug or need an extra field? Open an issue on the actor — fixes and new fields ship fast.
Legal & privacy
Data via the public Internet Archive Wayback Machine (availability + CDX APIs).
Keywords: wayback machine scraper, internet archive, web archive, snapshot history, cdx, osint, seo, domain history, JSON CSV Excel