Observed arrival · 2026-09-11
ProjectionBench asks whether models put feelings in your mouth
An independent benchmark and leaderboard for unsolicited emotional attribution by LLM assistants.
○Open
⊠Login
$Paid
†Ads
✦Pretty
●Pro
◎Niche
◉Human
⚑Risk
ƒJS
Field notes
The benchmark evaluates several kinds of affect attribution, including claims made after an explicit prohibition and assertions of hostility or tone. Its calibration scenario deliberately includes a user who genuinely expresses frustration, so the scoring is not designed to reward silence alone. The page also exposes versioned rubrics, raw metric caveats, and transcript-level findings, while noting that API and web surfaces should not be compared.
Observed signals
Read the marks
Editorial observations of this landing page, not a rating.
○OpenPublic substance visible
✦PrettyNotable craft visible
●ProPolished or operationally mature
◎NicheUnusually specific use
One card from the complete issue