Every major AI model ranked across benchmark quality, inference speed, agentic capability, programming aptitude, and cost efficiency — updated continuously from published evaluation data.
334
Tracked models
29
Providers
286
Benchmarked
13.1
Avg. index
334 models
| Rank | Model | Provider | Score | Benchmarks | Inference | Agentic | Programming | Value | Price |
|---|---|---|---|---|---|---|---|---|---|
| 81 | Claude Sonnet 4.6 claude-sonnet-4-6 multimodalvisionmulti-input reasoning | Anthropic | 12.6 Inference | 62.3 | 12.6 | 41.0 | 65.6 | 12.0 | $3 in / $15 out |
| 82 | GLM-5 glm-5 codeprogrammingtool use | Zhipu AI | 6.9 Inference | 0.0 | 6.9 | 35.1 | 60.4 | 40.0 | $1 in / $3.2 out |
| 83 | Qwen3 32B qwen3-32b textinference | Alibaba Cloud / Qwen Team | 2.2 Inference | 20.1 | 2.2 | 0.0 | 0.0 | 78.0 | $0.1 in / $0.3 out |
| 84 | ChatGPT-4o Latest chatgpt-4o-latest multimodalvisionmulti-input reasoning | OpenAI | 0.0 Inference | 53.0 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 85 | Claude 3.5 Haiku claude-3-5-haiku-20241022 codeprogrammingtool use | Anthropic | 0.0 Inference | 9.9 | 0.0 | 3.0 | 6.6 | 0.0 | |
| 86 | Claude 3.5 Sonnet claude-3-5-sonnet-20240620 multimodalvisionmulti-input reasoning | Anthropic | 0.0 Inference | 23.3 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 87 | Claude 3.5 Sonnet claude-3-5-sonnet-20241022 multimodalvisionmulti-input reasoning | Anthropic | 0.0 Inference | 32.0 | 0.0 | 38.7 | 11.1 | 0.0 | |
| 88 | Claude 3.7 Sonnet claude-3-7-sonnet-20250219 multimodalvisionmulti-input reasoning | Anthropic | 0.0 Inference | 42.3 | 0.0 | 49.1 | 37.7 | 0.0 | |
| 89 | Claude 3 Haiku claude-3-haiku-20240307 multimodalvisionmulti-input reasoning | Anthropic | 0.0 Inference | 5.3 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 90 | Claude 3 Opus claude-3-opus-20240229 multimodalvisionmulti-input reasoning | Anthropic | 0.0 Inference | 17.7 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 91 | Claude 3 Sonnet claude-3-sonnet-20240229 multimodalvisionmulti-input reasoning | Anthropic | 0.0 Inference | 9.2 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 92 | Claude Mythos Preview claude-mythos-preview multimodalvisionmulti-input reasoning | Anthropic | 0.0 Inference | 80.0 | 0.0 | 66.6 | 82.7 | 0.0 | |
| 93 | Claude Opus 4.1 claude-opus-4-1-20250805 multimodalvisionmulti-input reasoning | Anthropic | 0.0 Inference | 46.0 | 0.0 | 67.4 | 60.8 | 0.0 | |
| 94 | Claude Opus 4 claude-opus-4-20250514 multimodalvisionmulti-input reasoning | Anthropic | 0.0 Inference | 35.8 | 0.0 | 57.4 | 47.0 | 0.0 | |
| 95 | Claude Opus 4.5 claude-opus-4-5-20251101 multimodalvisionmulti-input reasoning | Anthropic | 0.0 Inference | 54.5 | 0.0 | 35.2 | 72.2 | 0.0 | |
| 96 | Claude Sonnet 4 claude-sonnet-4-20250514 multimodalvisionmulti-input reasoning | Anthropic | 0.0 Inference | 39.3 | 0.0 | 49.4 | 42.8 | 0.0 | |
| 97 | Codestral-22B codestral-22b textinference | Mistral AI | 0.0 Inference | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 98 | Command A+ command-a-plus-05-2026 multimodalvisionmulti-input reasoning | Cohere | 0.0 Inference | 35.2 | 0.0 | 0.0 | 15.3 | 0.0 | |
| 99 | Command R+ command-r-plus-04-2024 textinference | Cohere | 0.0 Inference | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 100 | DeepSeek-R1 deepseek-r1 textinference | DeepSeek | 0.0 Inference | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
Claude Sonnet 4.6
Anthropic
12.6
$3 in / $15 out
GLM-5
Zhipu AI
6.9
$1 in / $3.2 out
Qwen3 32B
Alibaba Cloud / Qwen Team
2.2
$0.1 in / $0.3 out
Want benchmark charts, model comparison, and pricing analytics?
Sign in to access the full interactive leaderboard with deep benchmark breakdowns and model comparison tools.
Open full leaderboardRankings are based on multi-dimensional evaluation across benchmark quality, inference efficiency, and cost-per-output. Scores are updated continuously and may differ from individual third-party benchmarks.
| N/A |
| N/A |
| N/A |
| N/A |
| N/A |
| N/A |
| N/A |
| N/A |
| N/A |
| N/A |
| N/A |
| N/A |
| N/A |
| N/A |
ChatGPT-4o Latest
OpenAI
0.0
N/A
Claude 3.5 Haiku
Anthropic
0.0
N/A
Claude 3.5 Sonnet
Anthropic
0.0
N/A
Claude 3.5 Sonnet
Anthropic
0.0
N/A
Claude 3.7 Sonnet
Anthropic
0.0
N/A
Claude 3 Haiku
Anthropic
0.0
N/A
Claude 3 Opus
Anthropic
0.0
N/A
Claude 3 Sonnet
Anthropic
0.0
N/A
Claude Mythos Preview
Anthropic
0.0
N/A
Claude Opus 4.1
Anthropic
0.0
N/A
Claude Opus 4
Anthropic
0.0
N/A
Claude Opus 4.5
Anthropic
0.0
N/A
Claude Sonnet 4
Anthropic
0.0
N/A
Codestral-22B
Mistral AI
0.0
N/A
Command A+
Cohere
0.0
N/A
Command R+
Cohere
0.0
N/A
DeepSeek-R1
DeepSeek
0.0
N/A