Skip to the card

Card 307 of 10002026-08-26 issue

Observed arrival · 2026-08-26

Fluency Bench: an AI test that tries to catch the bluff

fluencybench.com Observed source
Editorial interest 84/100 Selection signal · not a rating of the site

A 25-minute, three-round work sample that tests whether someone can brief an AI model, detect fabricated claims, and build a process that survives unseen cases.

Landing page captured for the 2026-08-26 issue.

Field notes

The assessment divides AI work into briefing, verification, and system-building rather than treating prompting as a single skill. Its published scoring includes six dimensions, a cost ceiling of $0.15 per item, and a visible overfit_delta comparing performance on seen and unseen cases. A sample report also shows adversarial results and category-level scores. The homepage says access codes travel by invite, while the rubric and sample evidence remain publicly inspectable.

Observed signals

Read the marks

Editorial observations of this landing page, not a rating.

OpenPublic substance visible
PrettyNotable craft visible
ProPolished or operationally mature
NicheUnusually specific use
ƒJavaScriptBrowser-side code central

One card from the complete issue

Eleven Views of a Dandelion

357,299 arrived 1,000 judged 1000 catalogued Enter the complete issue
fluencybench.com

Landing page observed 2026-08-26. The live site may have changed.