Observed arrival · 2026-09-27
Najd Arena puts Arabic AI models on a measurable field
An Arabic- and Saudi-focused AI benchmark compares model results across language, local knowledge, tool decisions, and other task areas.
○Open
⊠Login
$Paid
†Ads
✦Pretty
●Pro
◎Niche
◉Human
⚑Risk
ƒJS
- For
- Teams selecting Arabic-capable AI models
- Worth noticing
- Technical failures count as zero, and the site says score differences are descriptive rather than a significance claim.
Field notes
The interface separates direct execution from Pi agent runs and lets readers compare thinking settings rather than collapsing them into one headline score. Its answer-label breakdown distinguishes correct, possible-correct, partial, wrong, and technical-failure outcomes; the site says the underlying September 9 snapshot has not received a full-answer audit.
Observed signals
Read the marks
Editorial observations of this landing page, not a rating.
○OpenPublic substance visible
✦PrettyNotable craft visible
●ProPolished or operationally mature
◎NicheUnusually specific use
ƒJavaScriptBrowser-side code central
One card from the complete issue