Observed arrival · 2026-09-02
verityResearch freezes the verdict before the experiment
An independent research lab publishes pre-registered experiments on fine-tuning, tool use, and AI model behavior.
○Open
⊠Login
$Paid
†Ads
✦Pretty
●Pro
◎Niche
◉Human
⚑Risk
ƒJS
Field notes
The project uses a fixed sequence: define both arms, commit the interpretation and pass/fail rules, then read the evaluation output. Its sole listed study reports two evaluation sets—a held-out set of 106 prompts and an adversarial probe of 43—with answer rates as well as tool-call rates. The homepage links to a full write-up and GitHub code, and identifies the license as Apache-2.0.
Observed signals
Read the marks
Editorial observations of this landing page, not a rating.
○OpenPublic substance visible
✦PrettyNotable craft visible
●ProPolished or operationally mature
◎NicheUnusually specific use
◉HumanPersonal, local, civic, or handmade
One card from the complete issue