Every major AI model ranked across benchmark quality, inference speed, agentic capability, programming aptitude, and cost efficiency — updated continuously from published evaluation data.
334
Tracked models
29
Providers
286
Benchmarked
12.2
Avg. index
334 models
| Rank | Model | Provider | Score | Benchmarks | Inference | Agentic | Programming | Value | Price |
|---|---|---|---|---|---|---|---|---|---|
| 281 | Mistral Small 3.1 24B Instruct mistral-small-3.1-24b-instruct-2503 multimodalvisionmulti-input reasoning | Mistral AI | 0.0 Agentic | 15.1 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 282 | Mistral Small 3.2 24B Instruct mistral-small-3.2-24b-instruct-2506 multimodalvisionmulti-input reasoning | Mistral AI | 0.0 Agentic | 18.4 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 283 | Mistral Small 4 mistral-small-latest multimodalvisionmulti-input reasoning | Mistral AI | 0.0 Agentic | 31.3 | 23.2 | 0.0 | 0.0 | 81.7 | |
| 284 | North Mini Code 1.0 north-mini-code-1.0 codeprogrammingtool use | Cohere | 0.0 Agentic | 0.0 | 0.0 | 0.0 | 18.6 | 0.0 | N/A |
| 285 | Nova 2 Omni nova-2-omni multimodalvisionmulti-input reasoning | Amazon | 0.0 Agentic | 36.1 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 286 | Nova 2 Sonic nova-2-sonic multimodalvisionmulti-input reasoning | Amazon | 0.0 Agentic | 0.0 | 62.8 | 0.0 | 0.0 | 57.3 | $0.33 in / $2.75 out |
| 287 | Nova Lite nova-lite multimodalvisionmulti-input reasoning | Amazon | 0.0 Agentic | 12.9 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 288 | Nova Micro nova-micro textinference | Amazon | 0.0 Agentic | 8.4 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 289 | Nova Pro nova-pro multimodalvisionmulti-input reasoning | Amazon | 0.0 Agentic | 19.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 290 | Nemotron Nano 9B v2 nvidia-nemotron-nano-9b-v2 textinference | NVIDIA | 0.0 Agentic | 23.1 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 291 | o1-mini o1-mini textinference | OpenAI | 0.0 Agentic | 23.6 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 292 | o1-preview o1-preview codeprogrammingtool use | OpenAI | 0.0 Agentic | 40.2 | 0.0 | 0.0 | 8.1 | 0.0 | N/A |
| 293 | o1-pro o1-pro multimodalvisionmulti-input reasoning | OpenAI | 0.0 Agentic | 44.1 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 294 | o3-pro o3-pro-2025-06-10 multimodalvisionmulti-input reasoning | OpenAI | 0.0 Agentic | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 295 | Phi-3.5-mini-instruct phi-3.5-mini-instruct multimodalvisionmulti-input reasoning | Microsoft | 0.0 Agentic | 2.4 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 296 | Phi-3.5-MoE-instruct phi-3.5-moe-instruct multimodalvisionmulti-input reasoning | Microsoft | 0.0 Agentic | 7.5 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 297 | Phi-3.5-vision-instruct phi-3.5-vision-instruct multimodalvisionmulti-input reasoning | Microsoft | 0.0 Agentic | 2.3 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 298 | Phi 4 phi-4 textinference | Microsoft | 0.0 Agentic | 14.4 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 299 | Phi 4 Mini phi-4-mini textinference | Microsoft | 0.0 Agentic | 1.9 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 300 | Phi 4 Mini Reasoning phi-4-mini-reasoning textinference | Microsoft | 0.0 Agentic | 19.9 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
Mistral Small 3.1 24B Instruct
Mistral AI
0.0
N/A
Mistral Small 3.2 24B Instruct
Mistral AI
0.0
N/A
Mistral Small 4
Mistral AI
0.0
$0.15 in / $0.6 out
Want benchmark charts, model comparison, and pricing analytics?
Sign in to access the full interactive leaderboard with deep benchmark breakdowns and model comparison tools.
Open full leaderboardRankings are based on multi-dimensional evaluation across benchmark quality, inference efficiency, and cost-per-output. Scores are updated continuously and may differ from individual third-party benchmarks.
| N/A |
| $0.15 in / $0.6 out |
North Mini Code 1.0
Cohere
0.0
N/A
Nova 2 Omni
Amazon
0.0
N/A
Nova 2 Sonic
Amazon
0.0
$0.33 in / $2.75 out
Nova Lite
Amazon
0.0
N/A
Nova Micro
Amazon
0.0
N/A
Nova Pro
Amazon
0.0
N/A
Nemotron Nano 9B v2
NVIDIA
0.0
N/A
o1-mini
OpenAI
0.0
N/A
o1-preview
OpenAI
0.0
N/A
o1-pro
OpenAI
0.0
N/A
o3-pro
OpenAI
0.0
N/A
Phi-3.5-mini-instruct
Microsoft
0.0
N/A
Phi-3.5-MoE-instruct
Microsoft
0.0
N/A
Phi-3.5-vision-instruct
Microsoft
0.0
N/A
Phi 4
Microsoft
0.0
N/A
Phi 4 Mini
Microsoft
0.0
N/A
Phi 4 Mini Reasoning
Microsoft
0.0
N/A