Every major AI model ranked across benchmark quality, inference speed, agentic capability, programming aptitude, and cost efficiency — updated continuously from published evaluation data.
334
Tracked models
29
Providers
286
Benchmarked
29.3
Avg. index
334 models
| Rank | Model | Provider | Score | Benchmarks | Inference | Agentic | Programming | Value | Price |
|---|---|---|---|---|---|---|---|---|---|
| 321 | Ministral 3 (3B Instruct 2512) ministral-3-3b-instruct-2512 multimodalvisionmulti-input reasoning | Mistral AI | 0.0 overall | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 322 | Ministral 3 (8B Base 2512) ministral-3-8b-base-2512 multimodalvisionmulti-input reasoning | Mistral AI | 0.0 overall | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 323 | Ministral 3 (8B Instruct 2512) ministral-3-8b-instruct-2512 multimodalvisionmulti-input reasoning | Mistral AI | 0.0 overall | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 324 | Ministral 8B Instruct ministral-8b-instruct-2410 textinference | Mistral AI | 0.0 overall | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 325 | Mistral Large 2 mistral-large-2-2407 textinference | Mistral AI | 0.0 overall | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 326 | Mistral NeMo Instruct mistral-nemo-instruct-2407 textinference | Mistral AI | 0.0 overall | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 327 | Mistral Small mistral-small-2409 textinference | Mistral AI | 0.0 overall | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 328 | o3-pro o3-pro-2025-06-10 multimodalvisionmulti-input reasoning | OpenAI | 0.0 overall | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 329 | Qwen2.5-Coder 32B Instruct qwen-2.5-coder-32b-instruct textinference | Alibaba Cloud / Qwen Team | 0.0 overall | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 330 | Qwen2.5-Coder 7B Instruct qwen-2.5-coder-7b-instruct textinference | Alibaba Cloud / Qwen Team | 0.0 overall | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 331 | Qwen3.5-0.8B qwen3.5-0.8b multimodalvisionmulti-input reasoning | Alibaba Cloud / Qwen Team | 0.0 overall | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 332 | Qwen3-Coder qwen3-coder textinference | Alibaba Cloud / Qwen Team | 0.0 overall | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 333 | Qwen3-Next-80B-A3B-Base qwen3-next-80b-a3b-base textinference | Alibaba Cloud / Qwen Team | 0.0 overall | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 334 | U2 u2 textinference | Unisound | 0.0 overall | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
Ministral 3 (3B Instruct 2512)
Mistral AI
0.0
N/A
Ministral 3 (8B Base 2512)
Mistral AI
0.0
N/A
Ministral 3 (8B Instruct 2512)
Mistral AI
0.0
N/A
Page 17 of 17 · 334 models
Want benchmark charts, model comparison, and pricing analytics?
Sign in to access the full interactive leaderboard with deep benchmark breakdowns and model comparison tools.
Open full leaderboardRankings are based on multi-dimensional evaluation across benchmark quality, inference efficiency, and cost-per-output. Scores are updated continuously and may differ from individual third-party benchmarks.
| N/A |
| N/A |
Ministral 8B Instruct
Mistral AI
0.0
N/A
Mistral Large 2
Mistral AI
0.0
N/A
Mistral NeMo Instruct
Mistral AI
0.0
N/A
Mistral Small
Mistral AI
0.0
N/A
o3-pro
OpenAI
0.0
N/A
Qwen2.5-Coder 32B Instruct
Alibaba Cloud / Qwen Team
0.0
N/A
Qwen2.5-Coder 7B Instruct
Alibaba Cloud / Qwen Team
0.0
N/A
Qwen3.5-0.8B
Alibaba Cloud / Qwen Team
0.0
N/A
Qwen3-Coder
Alibaba Cloud / Qwen Team
0.0
N/A
Qwen3-Next-80B-A3B-Base
Alibaba Cloud / Qwen Team
0.0
N/A
U2
Unisound
0.0
N/A