Observed arrival · 2026-09-03
publicapidata measures what public-data extraction actually returns
A benchmark-driven collection of diagnostics, playbooks, and guides for extracting data from public APIs and websites.
Field notes
The project separates extraction reliability from extraction economics: its playbooks show both the per-item input cost and the proposed report, retainer, guide, or build revenue. Visible examples include a contact-address study across 400 domains and a transcript benchmark covering 150 videos, with results split by residential, datacenter, and unproxied egress. The pages also describe failure classes such as dead domains, crawler refusals, and sites that publish no contact information.
Observed signals
Read the marks
Editorial observations of this landing page, not a rating.
One card from the complete issue