Observed arrival · 2026-09-24
jextract: PDF fields with page coordinates
A document-extraction tool that turns PDFs into typed field values paired with their source spans and page coordinates.
○Open
⊠Login
$Paid
†Ads
✦Pretty
●Pro
◎Niche
◉Human
⚑Risk
ƒJS
- For
- Teams extracting structured fields from PDFs
- Worth noticing
- Preset taxonomies cover invoices, contracts, résumés, and purchase orders, with named fields shown for each.
Field notes
The page exposes preset field lists for invoices, contracts, résumés, and purchase orders. Its sample invoice has five chunks and shows separate locate and pick timings; the project reports that all 12 requested fields were found. The architecture description says parsing happens locally, while Jev makes the field-selection decisions, and long documents are windowed for parallel locating.
Observed signals
Read the marks
Editorial observations of this landing page, not a rating.
○OpenPublic substance visible
✦PrettyNotable craft visible
●ProPolished or operationally mature
◎NicheUnusually specific use
ƒJavaScriptBrowser-side code central
One card from the complete issue