Which AI model for which task — 2026
For every task: the strongest model, the best value for money and the best open-weights model — from the leaderboards with prices, updated daily.
| Task | Best | Best value | Open weights |
|---|---|---|---|
| Text | S Gemini 4 Argon (High)Google | S DeepSeek V4.1 Flash HFDeepSeek · $0.0015 / $0.09 | S MiMo-V2.6-Pro HFXiaomi · $0.40 / $0.80 |
| Coding | S Gemini 4 Argon (High)Google | S DeepSeek V4.1 Flash HFDeepSeek · $0.0015 / $0.09 | S MiMo-V2.6-Pro HFXiaomi · $0.40 / $0.80 |
| Web dev | S Claude Opus 5.5Anthropic · $4 / $20 | S DeepSeek V4.1 Flash HFDeepSeek · $0.0015 / $0.09 | S Kimi K3 HFMoonshot AI · $1 / $4 |
| Agents | S Claude Fable 5.1Anthropic · $10 / $50 | A DeepSeek V4.1 Flash HFDeepSeek · $0.0015 / $0.09 | A Kimi K3 HFMoonshot AI · $1 / $4 |
| Image | S gpt-image-2.5-sunburstOpenAI · $0.0361 | A Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)Google · $0.0336 | A qwen-image-2.1Alibaba |
| Video | S gemini-omni-1.1-flashGoogle · $1.50 / $9 | S Grok Imagine Video 1.5xAI · $0.08 | A minimax-h3MiniMax |
| Audio | S AuroraUnknown | S Eleven Multilingual v2ElevenLabs · $180 | — |
Best value: the cheapest model of the top tiers (S/A) at its cheapest provider. Prices per 1M tokens (input / output) unless stated otherwise.
Updated · Sources: LMArena (CC BY 4.0), LiteLLM (MIT), models.dev (MIT), TTS Arena
Arena (LMArena) scores: Arena Leaderboard Dataset by LMArena, CC BY 4.0. Modified by fedi.software (filtered, re-ranked, aggregated). Not endorsed by LMArena.
Speech ratings: TTS Arena V2 by TTS-AGI (crowdsourced blind votes), used with permission — keep a link to the arena. Modified by fedi.software (filtered, tiered). Not endorsed by TTS-AGI.
Embed on your site
Paste this code where the infographic should appear: it updates by itself. Free to use — keep the attribution link.
One row per task
Each row of this chart is one of our leaderboards: text, coding, web development, agents, image and video. For every task you get three picks instead of a single winner, because the strongest model is rarely the one most people should pay for, and many readers want something they can download.
What the three columns mean
- Best — first place on that board. Ratings come from LMArena, where people compare two anonymous answers and vote.
- Best value — the cheapest model with a price among tiers S and A, measured by the blended price at its cheapest provider. If no S or A model has a price, the pick moves down to tier B.
- Open weights — the highest-rated model whose weights you can download and run on your own hardware.
Under each name you see the tier letter, the vendor and the price. Text boards are priced per 1M tokens of input and output; image models per picture and video per second. A click on a name opens the model in the side-by-side comparison, where you can put it next to the other two picks.
How to use it without being misled
First place is less decisive than it looks. Our tiers group models whose confidence intervals overlap, so the second and third model of a board are often statistically level with the leader and may be much cheaper. A rating also measures what arena voters preferred on their own prompts — mostly in English — not accuracy on your documents or your codebase. Treat the chart as a short list, then open the full board, read the votes and the interval, and try two or three candidates on a real task.
The open-weights pick is the starting point for running a model locally. Before you download it, check which open model fits your hardware: the memory it needs depends on its size and on the quantization you choose.
Yearly editions
Leaders change every few weeks, so the chart keeps one edition per year. Open a past year to see which models were new then and which ones have dropped out since.