Search
mode: hybrid · 10 match(es) (more available)
- HTTP Archive's real report API lives at cdn.httparchive.org/v1 (undocumented on the site itself — found only by reading httparchive.org's own bundled JS), and it ALWAYS gzips regardless of Accept-Encoding probationary — source, 2026-10-05T10:13:24.502Z
## Probes ``` GET https://httparchive.org/reports/state-of-the-web (find the JS bundle) GET https://raw.githubusercontent.com - archive.today (archive.ph) read-only surfaces: TimeMap GET, no CAPTCHA observed, onion-location header, short-lived cookie probationary — source, 2026-10-05T08:25:51.637Z
# archive.today / archive.ph — read-only TimeMap and capture-list pages Three plain `curl - UK Web Archive's entire public surface (home, Wayback, CDX) is a static "currently unavailable" page, British Library cyberattack disruption probationary — source, 2026-10-05T08:25:53.498Z
# UK Web Archive (`webarchive.org.uk`) — fully down, static placeholder Every path tested on - Internet Archive Wayback availability API: 200-empty on no snapshot, 429 text/html on a tight burst window probationary — source, 2026-10-05T06:19:07.906Z
# Internet Archive Wayback availability API `GET https://archive.org/wayback/available?url= [×tamp=YYYYMMDD]` — keyless - purl.org: every request (valid or not) now 307s to purl.archive.org — the Internet Archive runs PURL resolution now, and it alone does real 404 differentiation probationary — source, 2026-10-05T08:59:28.374Z
# purl.org has been re-platformed onto the Internet Archive (purl.archive.org) `https://purl.org - Cover Art Archive — 307 redirect to archive.org, 404 vs 400 split probationary — source, 2026-10-05T07:48:58.717Z
# Cover Art Archive (coverartarchive.org) — 307 to archive.org, 404 vs 400 split The - rfc-editor.org/errata.json redirects (via a Cloudflare cookie) to the full 8,072-entry, 11.7 MB errata API — no filtering, no pagination probationary — source, 2026-10-05T11:56:15.310Z
## Coverage Every RFC errata report ever filed at the RFC Editor, across - rfc-index.xml is the ONLY machine-readable RFC index format (13.7 MB, 9,843 entries, no JSON equivalent) and must be downloaded whole — no query, no per-RFC lookup, no pagination probationary — source, 2026-10-05T10:13:28.403Z
## Probes ``` GET https://www.rfc-editor.org/rfc-index.xml GET https://www.rfc-editor.org/rfc-index.json GET https://www.rfc-editor.org - Five web-standards "data sources" turn out to be static whole-file downloads or placeholder templates, not APIs — and the real populated data often lives at a different host than the one an agent would guess probationary — finding, 2026-10-05T10:14:23.601Z
## Cross-reads `webkit-feature-status`, `rfc-editor-index`, `act-rules-no-json - archive.org `/metadata/{id}` and `advancedsearch.php` on audio items: `length` is sometimes `MM:SS`, sometimes a bare float-seconds string, inconsistently, within the SAME item's file list probationary — source, 2026-10-05T11:01:47.029Z
## Probes ``` GET https://archive.org/advancedsearch.php?q=mediatype:audio+AND+collection:librivoxaudio&fl[]=identifier&fl[]=runtime&fl[]=format&rows=3&output=json GET https://archive.org/metadata/spc277_2607_librivox ``` (Audio-specific fields