Skip to the card

Card 750 of 10002026-09-11 issue

Observed arrival · 2026-09-11

Sane Labs, Teaching Small Models to Read

sanelabs.org Observed source
Editorial interest 86/100 Selection signal · not a rating of the site

An independent language-model project training small models from scratch, with knowledge intended to come from supplied context rather than memorized weights.

Landing page captured for the 2026-09-11 issue.

Field notes

The project exposes unusually detailed training records, including model shape, token counts, hardware, precision, and weight size. Sane-47M and Sane-118M use an in-house BPE tokenizer and 1,024-token context, while the in-training Synth-2 expands to a listed 8k–32k context and uses 12 experts plus a shared expert. The page also separates released work from private or unfinished experiments rather than presenting every project as available.

Observed signals

Read the marks

Editorial observations of this landing page, not a rating.

OpenPublic substance visible
PrettyNotable craft visible
ProPolished or operationally mature
NicheUnusually specific use

One card from the complete issue

Projection by Pedal

381,023 arrived 1,000 judged 1000 catalogued Enter the complete issue
sanelabs.org

Landing page observed 2026-09-11. The live site may have changed.