Operated by Sourcelane✓ verified
A scraper that enumerates Internet Archive (Wayback Machine) snapshots for a URL, page path, hostname, or entire domain (including subdomains), producing per-capture metadata such as capture timestamp, direct snapshot URL, HTTP status code, MIME type, byte size, and a content digest for deduplication; supports match modes (exact URL, prefix, host, domain), date-range / status-code / MIME-type filtering, collapse/merge of identical captures by content digest, and cursor-based pagination using the Wayback Machine CDX resume key to retrieve large or unlimited numbers of captures across multiple runs—outputs structured, snapshot-level archival history and change-tracking metadata for each archived capture.
Verified Sep 28, 3:55 AM
$0.001
per archived snapshot
Up to 100 per call. Only pay for results returned; failed calls are refunded.
| Window | Uptime | Success rate | p50 | p95 | Calls |
|---|---|---|---|---|---|
| 24h | — | — | — | — | 0 |
| 7d | — | — | — | — | 0 |
| 30d | — | — | — | — | 0 |
Call it through Sourcelane's gateway or MCP server. We run the connector, apply your agent's spend guardrails, and bill only the results returned.
Call it via Sourcelane
curl -X POST https://api.usesourcelane.com/v1/call \
-H "Authorization: Bearer sl_live_your_agent_key" \
-H "Content-Type: application/json" \
-d '{"listing":"wayback-machine-snapshots-scraper","params":{"url":"example.com","maxItems":5},"maxResults":10}'Request params
| Field | Type | Required | Description |
|---|---|---|---|
url | string | required | The URL to look up. Use 'Match Type = Prefix' or 'Domain' to scan an entire site rather than a single page. |
matchType | string | optional | How to match the URL. 'Exact' only returns snapshots of the exact URL; 'Domain' returns snapshots for the domain and all subdomains. |
dateFrom | string | optional | Optional. Only return snapshots archived on or after this UTC date. |
dateTo | string | optional | Optional. Only return snapshots archived on or before this UTC date. |
statusCode | string | optional | Optional. Only return snapshots with this HTTP status (e.g. 200, 301, 404). |
mimeType | string | optional | Optional. Only return snapshots with this MIME type (e.g. text/html, image/jpeg). |
collapseDuplicates | boolean | optional | If on, identical captures (same content digest) are deduplicated so you only see snapshots where the content actually changed. |
maxItems | integer | optional | Maximum snapshots to return. 0 = no cap (paginates until everything is fetched or the run times out). Capped at 100 per call. |
pageId | string | optional | Optional. Paste the NEXT_PAGE_ID (CDX resume key) from the previous run's Key-value store to continue from where you left off. |
Each result contains
| Field | Type | Description |
|---|---|---|
timestamp | string | Timestamp |
archivedAt | string | Archived At |
originalUrl | string | Original Url |
snapshotUrl | string | Snapshot Url |
statusCode | integer | Status Code |
mimeType | string | Mime Type |
contentLength | integer | Content Length |
digest | string | Digest |
Pick your tool and connect in under a minute. Then just ask — for example: “For these 50 domains, find contact emails and what e-commerce platform they run on.”
Connect Claude
Web, desktop and mobile. Paste one URL.
Adds the Sourcelane MCP server with your key as a header.
claude mcp add --transport http sourcelane https://api.usesourcelane.com/mcp \
--header "Authorization: Bearer sl_live_your_agent_key"Alternative to the connector URL: add to claude_desktop_config.json, then restart Claude.
{
"mcpServers": {
"sourcelane": {
"command": "npx",
"args": [
"-y",
"mcp-remote",
"https://api.usesourcelane.com/mcp",
"--header",
"Authorization:${AUTH_HEADER}"
],
"env": {
"AUTH_HEADER": "Bearer sl_live_your_agent_key"
}
}
}
}Add to ~/.cursor/mcp.json (or Windsurf's mcp_config.json).
{
"mcpServers": {
"sourcelane": {
"url": "https://api.usesourcelane.com/mcp",
"headers": {
"Authorization": "Bearer sl_live_your_agent_key"
}
}
}
}Save as .vscode/mcp.json. VS Code prompts for your key once.
{
"servers": {
"sourcelane": {
"type": "http",
"url": "https://api.usesourcelane.com/mcp",
"headers": {
"Authorization": "Bearer ${input:sourcelane-key}"
}
}
},
"inputs": [
{
"type": "promptString",
"id": "sourcelane-key",
"description": "Sourcelane agent key",
"password": true
}
]
}Create a GPT → Actions → Import from URL, then Authentication: API Key, Bearer.
https://usesourcelane.com/openapi.jsonOne POST. Pass maxResults to cap cost.
curl -X POST https://api.usesourcelane.com/v1/call \
-H "Authorization: Bearer sl_live_your_agent_key" \
-H "Content-Type: application/json" \
-d '{"listing":"wayback-machine-snapshots-scraper","params":{"url":"example.com","maxItems":5},"maxResults":10}'Wrap the REST call as a tool, or point an MCP client at the endpoint.
import requests
res = requests.post(
"https://api.usesourcelane.com/v1/call",
headers={"Authorization": "Bearer sl_live_your_agent_key"},
json={
"listing": "wayback-machine-snapshots-scraper",
"params": {"url":"example.com","maxItems":5},
"maxResults": 10,
},
timeout=300,
)
body = res.json()
print(body["receipt"]["chargedMicros"], "micro-USD for", body["receipt"]["results"], "results")
print(body["data"])—
0 reviews
Leave a review
Loading reviews…