Observed arrival · 2026-09-25
WebGraph, the honest web reader
An open-source, locally run tool that reads websites and returns their public pages as Markdown, with notes on how each page was obtained.
○Open
⊠Login
$Paid
†Ads
✦Pretty
●Pro
◎Niche
◉Human
⚑Risk
ƒJS
- For
- Developers building website readers or crawlers
- Worth noticing
- The example crawl labels login redirects, 503s, bot challenges, hidden text, and off-screen links as separate outcomes.
Field notes
The described discovery queue starts with robots.txt and sitemap.xml, then follows links rather than guessing a site’s page list. Each page is fetched in two ways—a plain request and a Chromium render—and the output records where content came from or why a page was refused. One reported university crawl queued 17,000 URLs, a third PDFs that were counted but not fetched.
Observed signals
Read the marks
Editorial observations of this landing page, not a rating.
○OpenPublic substance visible
✦PrettyNotable craft visible
●ProPolished or operationally mature
◎NicheUnusually specific use
ƒJavaScriptBrowser-side code central
One card from the complete issue