General
Coding
Agents
Math
Vision
Search
Writing
Language: English
Live data for 2026-10 (recalculated daily)
· Methodology v1
· Methodology
· History
· JSON
#
Tier
Model
fedi-index
Percentile
People prefer
Solves tasks
Value
Sources
Score decomposition
13 B Kimi K3 Moonshot AI · Open weights 1,456 ±5 79 1,459 1,449 +44 A τ ›
14 B GLM 5.2 Zhipu AI · Open weights 1,454 ±5 77 1,456 1,449 +45 A τ ›
17 B DeepSeek V4.1 Flashprovisional DeepSeek · Open weights 1,449 ±2 72 1,459 — +92◆ A ›
19 B Hy4 previewprovisional Tencent · Open weights 1,448 ±3 68 1,458 — — A ›
24 B GLM 5.3provisional Zhipu AI · Open weights 1,442 ±3 60 1,451 — +60 A ›
27 B DeepSeek V4 Pro 0423provisional DeepSeek · Open weights 1,435 ±7 54 1,445 — +33 A ›
30 C GLM 5.3 Flashprovisional Zhipu AI · Open weights 1,430 ±2 49 1,439 — +105◆ A ›
32 C Qwen3.8 Flash Nextprovisional Alibaba · Open weights 1,428 ±3 46 1,437 — — A ›
33 C MiMo-V2.6-Flashprovisional Xiaomi · Open weights 1,428 ±6 44 1,438 — +98◆ A ›
36 C Qwen3.8 27Bprovisional Alibaba · Open weights 1,424 ±3 39 1,433 — +46 A ›
42 C Hy3provisional Tencent · Open weights 1,408 ±5 28 1,416 — +68 A ›
47 D MiniMax M3provisional MiniMax · Open weights 1,402 ±4 19 1,410 — +50 A ›
49 D MiMo-V2.5-Proprovisional Xiaomi · Open weights 1,399 ±4 16 1,407 — +80 A ›
53 D GLM 5provisional Zhipu AI · Open weights 1,388 ±8 7 — 1,411 +23 τ ›
54 D Qwen3.5 397B A17Bprovisional Alibaba · Open weights 1,388 ±8 7 — 1,411 +27 τ ›
55 D Inkling Smallprovisional Thinking Machines Lab · Open weights 1,386 ±6 5 1,394 — +38 A ›
56 D Inklingprovisional Thinking Machines Lab · Open weights 1,384 ±4 4 1,392 — +27 A ›
57 D Mistral Medium 3.5provisional Mistral AI · Open weights 1,377 ±6 2 1,384 — -19 A ›
…
A Arena (human votes)τ τ²-bench (agents) ◆ Pareto frontier: nothing is both better and cheaper
fedi-index is an Elo-equivalent on a fixed scale: comparable between months within one methodology version. Tiers S–D: by percentile in the slice, and no better than the confidence interval allows (cut-offs in the methodology). “Provisional” — only one source family measured the model here.
Methodology
Embed on your site
The table as a widget for your site: it updates by itself and keeps the attribution of the sources.
Language
English
Español
Français
Português (Brasil)
Deutsch
Italiano
繁體中文
Русский
日本語
Bahasa Indonesia
한국어
العربية
ไทย
Türkçe
Polski
Deutsch (Schweiz)
Norsk
Slovenščina
Dansk
Svenska
Nederlands
Ελληνικά
Suomi
Română
Theme
Automatic
Light
Dark
Height, px
Code
Copy code
Copied
JSON
fedi-index © fedi.software, CC BY 4.0; components keep the licences of their sources.
fedi-index © fedi.software, CC BY 4.0; components keep the licences of their sources. · Sources:
LMArena (CC BY 4.0),
Epoch AI (CC BY 4.0),
τ²-bench (Sierra) (MIT),
models.dev (MIT),
LiteLLM (MIT)
Arena Leaderboard Dataset by LMArena, CC BY 4.0. Modified by fedi.software (equated, weighted, aggregated). Not endorsed by LMArena.
Epoch AI, 'Capabilities & benchmarking'. Published online at epoch.ai. Retrieved from 'https://epoch.ai/benchmarks'. CC BY. ECI by Epoch AI; Epoch-run results only. Modified by fedi.software.
τ²-bench leaderboard, © Sierra Research, MIT License (Yao et al. 2024, arXiv:2406.12045; Barres et al. 2025, arXiv:2506.07982). Sierra-run text submissions only.
Pricing: models.dev, MIT License, © 2025 models.dev.
Pricing: LiteLLM model_prices_and_context_window.json, MIT License, © 2023 Berri AI.