Every major AI model ranked across benchmark quality, inference speed, agentic capability, programming aptitude, and cost efficiency — updated continuously from published evaluation data.
334
Tracked models
29
Providers
286
Benchmarked
28.3
Avg. index
334 models
| Rank | Model | Provider | Score | Benchmarks | Inference | Agentic | Programming | Value | Price |
|---|---|---|---|---|---|---|---|---|---|
| 321 | Ministral 8B Instruct ministral-8b-instruct-2410 textinference | Mistral AI | 0.0 Benchmarks | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 322 | Mistral Large 2 mistral-large-2-2407 textinference | Mistral AI | 0.0 Benchmarks | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 323 | Mistral NeMo Instruct mistral-nemo-instruct-2407 textinference | Mistral AI | 0.0 Benchmarks | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 324 | Mistral Small mistral-small-2409 textinference | Mistral AI | 0.0 Benchmarks | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 325 | North Mini Code 1.0 north-mini-code-1.0 codeprogrammingtool use | Cohere | 0.0 Benchmarks | 0.0 | 0.0 | 0.0 | 18.6 | 0.0 | N/A |
| 326 | Nova 2 Sonic nova-2-sonic multimodalvisionmulti-input reasoning | Amazon | 0.0 Benchmarks | 0.0 | 62.8 | 0.0 | 0.0 | 57.3 | $0.33 in / $2.75 out |
| 327 | o3-pro o3-pro-2025-06-10 multimodalvisionmulti-input reasoning | OpenAI | 0.0 Benchmarks | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 328 | Qwen2.5-Coder 32B Instruct qwen-2.5-coder-32b-instruct textinference | Alibaba Cloud / Qwen Team | 0.0 Benchmarks | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 329 | Qwen2.5-Coder 7B Instruct qwen-2.5-coder-7b-instruct textinference | Alibaba Cloud / Qwen Team | 0.0 Benchmarks | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 330 | Qwen3.5-0.8B qwen3.5-0.8b multimodalvisionmulti-input reasoning | Alibaba Cloud / Qwen Team | 0.0 Benchmarks | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 331 | Qwen3-Coder qwen3-coder textinference | Alibaba Cloud / Qwen Team | 0.0 Benchmarks | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 332 | Qwen3-Coder 480B A35B Instruct qwen3-coder-480b-a35b-instruct codeprogrammingtool use | Alibaba Cloud / Qwen Team | 0.0 Benchmarks | 0.0 | 0.0 | 50.7 | 32.1 | 0.0 | |
| 333 | Qwen3-Next-80B-A3B-Base qwen3-next-80b-a3b-base textinference | Alibaba Cloud / Qwen Team | 0.0 Benchmarks | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 334 | U2 u2 textinference | Unisound | 0.0 Benchmarks | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
Ministral 8B Instruct
Mistral AI
0.0
N/A
Mistral Large 2
Mistral AI
0.0
N/A
Mistral NeMo Instruct
Mistral AI
0.0
N/A
Page 17 of 17 · 334 models
Want benchmark charts, model comparison, and pricing analytics?
Sign in to access the full interactive leaderboard with deep benchmark breakdowns and model comparison tools.
Open full leaderboardRankings are based on multi-dimensional evaluation across benchmark quality, inference efficiency, and cost-per-output. Scores are updated continuously and may differ from individual third-party benchmarks.
| N/A |
Mistral Small
Mistral AI
0.0
N/A
North Mini Code 1.0
Cohere
0.0
N/A
Nova 2 Sonic
Amazon
0.0
$0.33 in / $2.75 out
o3-pro
OpenAI
0.0
N/A
Qwen2.5-Coder 32B Instruct
Alibaba Cloud / Qwen Team
0.0
N/A
Qwen2.5-Coder 7B Instruct
Alibaba Cloud / Qwen Team
0.0
N/A
Qwen3.5-0.8B
Alibaba Cloud / Qwen Team
0.0
N/A
Qwen3-Coder
Alibaba Cloud / Qwen Team
0.0
N/A
Qwen3-Coder 480B A35B Instruct
Alibaba Cloud / Qwen Team
0.0
N/A
Qwen3-Next-80B-A3B-Base
Alibaba Cloud / Qwen Team
0.0
N/A
U2
Unisound
0.0
N/A