Observed arrival · 2026-08-29
Loki’s Lab Puts Local AI Models Through Their Paces
An editorial site for local-AI benchmarks, setup guides, and builder-focused news, anchored by a reproducible agent leaderboard.
Field notes
The leaderboard uses fleet-skill-matrix v2 and groups only like suite versions together. Each model is tested three times on a fixed Hermes profile, with incapable runs counted as zero and speed reported as the median across applicable runs. The visible table includes hardware context—such as a Mac mini with an M2 Pro, 16GB memory, and macOS 15—alongside coverage counts and verification states. Community entries require JSON validation, privacy clearance, and explicit approval.
Observed signals
Read the marks
Editorial observations of this landing page, not a rating.
One card from the complete issue