Wayback Machine
The Internet Archive's Wayback Machine: the archived capture of a URL closest to a date.
- Provider id:
wayback - Tools: 1, at $0.001 per call
- Categories: Content Extraction, Research & Papers
- Homepage: web.archive.org
- Provider docs: archive.org/help/wayback_api.php
- Attribution: Internet Archive Wayback Machine
| Tool | Price | What it does |
|---|---|---|
wayback/snapshot | $0.001 | Find the Internet Archive's capture of a URL closest to a date (or its newest capture). |
Inputs are JSON bodies, and unknown fields are rejected with 422 invalid_input (not charged). POST /v1/inspect with a tool's id returns the same input schema, plus the output schema.
Wayback Machine Snapshot
wayback/snapshot · $0.001 per call (local) · Content Extraction, Research & Papers
Find the Internet Archive's capture of a URL closest to a date (or its newest capture).
Asks the Wayback Machine for the archived copy of a page nearest to a moment you give (or the newest one) and returns the web.archive.org link, the capture time and the HTTP status the site returned when it was archived. Use it for 'what did this page say in 2019', dead links and changed pricing pages. It returns the link, not the page: read the archived_url with firecrawl/scrape or jina/read. The Wayback Machine allows about 15 lookups a minute, so results are cached.
Related: firecrawl/scrape, jina/read, serper/search.
Input
| Field | Type | Required | Description |
|---|---|---|---|
url | string | yes | The page to look up, e.g. 'https://www.python.org/downloads' or 'python.org'. 3–2000 characters. Pattern ^\S+$. |
at | string | no | Closest to this moment: '2015', '2015-06-30' or '2015-06-30T12:00:00Z' (default: the newest capture). |
Example
{
"url": "https://www.python.org/",
"at": "2010-01-01"
}node akashi.mjs run wayback/snapshot --input '{"url":"https://www.python.org/","at":"2010-01-01"}'