Compare and rank AI models by quality score, cost, speed, and context window. Interactive chart with 30 models. Find the best model for your workload and budget.
AI model comparison dashboard with 30 models ranked by quality, cost, speed, and context window. Filter by task, tier, and sort by the metrics that matter for your workload.
| # | Model | Tier | Quality | Price (In/Out) | Speed | Context | Value Score |
|---|---|---|---|---|---|---|---|
| 1 | Claude Opus 5 Anthropic | Frontier | 61 | $5.00 / $25.00 | 52 tok/s | 1.0M | 2 |
| 2 | Claude Fable 5 Anthropic | Frontier | 60 | $10.00 / $50.00 | 71 tok/s | 1.0M | 1 |
| 3 | GPT-5.6 Sol OpenAI | Frontier | 59 | $5.00 / $30.00 | 85 tok/s | 1.1M | 2 |
| 4 | Kimi K3 Moonshot AI | Frontier | 57 | $3.00 / $15.00 | 62 tok/s | 1.0M | 3 |
| 5 | Claude Opus 4.8 Anthropic | Frontier | 56 | $5.00 / $25.00 | 30 tok/s | 1.0M | 2 |
| 6 | GPT-5.6 Terra OpenAI | Premium | 55 | $2.00 / $12.00 | 75 tok/s | 1.1M | 4 |
| 7 | GPT-5.5 OpenAI | Premium | 55 | $5.00 / $30.00 | 80 tok/s | 1.1M | 2 |
| 8 | Claude Opus 4.7 Anthropic | Premium | 54 | $5.00 / $25.00 | 48 tok/s | 1.0M | 2 |
| 9 | Grok 4.5 xAI | Premium | 54 | $2.00 / $6.00 | 70 tok/s | 500K | 7 |
| 10 | Claude Sonnet 5 Anthropic | Premium | 53 | $2.00 / $10.00 | 78 tok/s | 1.0M | 4 |
| 11 | GPT-5.4 OpenAI | Mid-Range | 51 | $2.50 / $15.00 | 116 tok/s | 1.1M | 3 |
| 12 | GPT-5.6 Luna OpenAI | Budget | 51 | $0.20 / $1.20 | 150 tok/s | 1.1M | 36 |
| 13 | GLM-5.2 Z AI | Mid-Range | 51 | $0.50 / $2.00 | 170 tok/s | 128K | 20 |
| 14 | Gemini 3.6 Flash Google | Mid-Range | 50 | $1.50 / $7.50 | 251 tok/s | 1.0M | 6 |
| 15 | Gemini 3.5 Flash Google | Budget | 50 | $1.50 / $9.00 | 178 tok/s | 1.0M | 5 |
| 16 | Gemini 3.1 Pro Google | Frontier | 47 | $2.50 / $15.00 | 113 tok/s | 1.0M | 3 |
| 17 | Claude Sonnet 4.6 Anthropic | Mid-Range | 47 | $3.00 / $15.00 | 46 tok/s | 1.0M | 3 |
| 18 | Gemini 3 Flash Google | Budget | 46 | $0.07 / $0.30 | 160 tok/s | 1.0M | 123 |
| 19 | DeepSeek V4 Pro DeepSeek | Mid-Range | 44 | $0.43 / $0.87 | 71 tok/s | 128K | 34 |
| 20 | KAT-Coder-Pro V2 KwaiPilot | Mid-Range | 44 | $0.30 / $1.20 | 100 tok/s | 256K | 29 |
| 21 | MiniMax M3 MiniMax | Mid-Range | 44 | $0.30 / $1.20 | 93 tok/s | 205K | 29 |
| 22 | Grok 4 xAI | Premium | 43 | $3.00 / $15.00 | 66 tok/s | 2.0M | 2 |
| 23 | Gemini 3 Pro Google | Premium | 40 | $2.00 / $12.00 | 90 tok/s | 1.0M | 3 |
| 24 | DeepSeek V4 Flash DeepSeek | Budget | 40 | $0.14 / $0.28 | 118 tok/s | 128K | 95 |
| 25 | MiniMax M2.7 MiniMax | Mid-Range | 38 | $0.30 / $1.20 | 51 tok/s | 205K | 25 |
| 26 | GPT-4o Mini OpenAI | Budget | 38 | $0.15 / $0.60 | 180 tok/s | 128K | 51 |
| 27 | Claude Haiku 4.5 Anthropic | Budget | 37 | $1.00 / $5.00 | 95 tok/s | 200K | 6 |
| 28 | DeepSeek R1 (Free) DeepSeek | Free | 27 | Free / Free | 40 tok/s | 64K | 270 |
| 29 | Qwen 2.5 VL 72B (Free) Alibaba | Free | 15 | Free / Free | 50 tok/s | 128K | 150 |
| 30 | Llama 3.3 70B (Free) Meta | Free | 14 | Free / Free | 80 tok/s | 128K | 140 |
Was this tool helpful?
An AI model comparison ranks language models side by side on the metrics that affect your costs and output quality: intelligence score, price per million tokens, tokens per second, and context window size.
This dashboard pulls scores from the Artificial Analysis Intelligence Index, pricing from OpenRouter and provider APIs, and speed data from independent benchmarks. Filter by task (coding, writing, analysis) or price tier to narrow results to models that fit your workload.
The Value Score divides quality by cost. Higher means more intelligence per dollar. Use it to find models that punch above their price. For a ranked view with editorial analysis, see the LLM Leaderboard.
Free AI optimization and data conversion tools.