Skip to the card

Card 430 of 9752026-09-24 issue

Observed arrival · 2026-09-24

jeval: evaluator judgments in one API call

jeval.dev Visit website
Editorial interest 77/100 Selection signal · not a rating of the site

jeval presents an API that runs evaluators from tools such as RAGAS and promptfoo against one submitted item, returning scores, probabilities, and confidence.

Landing page captured for the 2026-09-24 issue.
For
Developers evaluating LLM and RAG outputs
Worth noticing
The page describes per-chunk grounding verdicts and per-step or per-call agent judgments alongside standard evaluator scores.

Field notes

The page describes a single request carrying an item’s input, output, expected answer, context, criteria, tool calls, or trajectory, with selected evaluator questions run against that state. Its examples extend beyond one overall score: retrieved chunks receive individual grounding verdicts, while agent trajectories can be judged step by step or call by call. The site reports 43 evaluators, with 33 answered by Jev and 10 in code; its latency and cost figures are homepage claims.

Observed signals

Read the marks

Editorial observations of this landing page, not a rating.

○OpenPublic substance visible
✦PrettyNotable craft visible
◎NicheUnusually specific use
ƒJavaScriptBrowser-side code central

One card from the complete issue

The Busker’s Compute Bill

314,149 arrived 1,000 judged 975 catalogued Enter the complete issue
jeval.dev

Landing page observed 2026-09-24. The live site may have changed.