Every major AI model ranked across benchmark quality, inference speed, agentic capability, programming aptitude, and cost efficiency — updated continuously from published evaluation data.
334
Tracked models
29
Providers
286
Benchmarked
12.2
Avg. index
334 models
| Rank | Model | Provider | Score | Benchmarks | Inference | Agentic | Programming | Value | Price |
|---|---|---|---|---|---|---|---|---|---|
| 241 | Llama 3.2 11B Instruct llama-3.2-11b-instruct multimodalvisionmulti-input reasoning | Meta | 0.0 Agentic | 3.8 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 242 | Llama 3.2 3B Instruct llama-3.2-3b-instruct textinference | Meta | 0.0 Agentic | 4.8 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 243 | Llama 3.2 90B Instruct llama-3.2-90b-instruct multimodalvisionmulti-input reasoning | Meta | 0.0 Agentic | 14.9 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 244 | Llama 3.3 70B Instruct llama-3.3-70b-instruct textinference | Meta | 0.0 Agentic | 18.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 245 | Llama-3.3 Nemotron Super 49B v1 llama-3.3-nemotron-super-49b-v1 textinference | NVIDIA | 0.0 Agentic | 21.3 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 246 | Llama 4 Maverick llama-4-maverick multimodalvisionmulti-input reasoning | Meta | 0.0 Agentic | 32.6 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 247 | Llama 4 Scout llama-4-scout multimodalvisionmulti-input reasoning | Meta | 0.0 Agentic | 27.6 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 248 | LongCat-Flash-Thinking longcat-flash-thinking codeprogrammingtool use | Meituan | 0.0 Agentic | 48.2 | 0.0 | 0.0 | 18.4 | 0.0 | |
| 249 | Magistral Medium magistral-medium multimodalvisionmulti-input reasoning | Mistral AI | 0.0 Agentic | 20.6 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 250 | Magistral Small 2506 magistral-small-2506 textinference | Mistral AI | 0.0 Agentic | 22.7 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 251 | MAI-Code-1-Flash mai-code-1-flash codeprogrammingtool use | Microsoft | 0.0 Agentic | 32.2 | 0.0 | 0.0 | 21.4 | 0.0 | N/A |
| 252 | MAI-Thinking-1 mai-thinking-1 codeprogrammingtool use | Microsoft | 0.0 Agentic | 60.1 | 0.0 | 0.0 | 32.2 | 0.0 | N/A |
| 253 | MedGemma 4B IT medgemma-4b-it multimodalvisionmulti-input reasoning | Google | 0.0 Agentic | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 254 | Mercury 2 mercury-2 codeprogrammingtool use | Inception | 0.0 Agentic | 42.7 | 68.9 | 0.0 | 15.3 | 84.4 | $0.25 in / $0.75 out |
| 255 | MiMo-V2.5 mimo-v2.5 multimodalvisionmulti-input reasoning | Xiaomi | 0.0 Agentic | 47.7 | 84.8 | 0.0 | 27.2 | 92.7 | $0.168 in / $0.336 out |
| 256 | MiMo-V2.5-Pro mimo-v2.5-pro codeprogrammingtool use | Xiaomi | 0.0 Agentic | 36.2 | 84.8 | 0.0 | 56.3 | 78.0 | $0.435 in / $0.87 out |
| 257 | MiMo-V2-Omni mimo-v2-omni multimodalvisionmulti-input reasoning | Xiaomi | 0.0 Agentic | 0.0 | 0.0 | 0.0 | 50.9 | 0.0 | N/A |
| 258 | MiMo-V2-Pro mimo-v2-pro codeprogrammingtool use | Xiaomi | 0.0 Agentic | 0.0 | 0.0 | 0.0 | 61.9 | 0.0 | N/A |
| 259 | MiniCPM-SALA minicpm-sala textinference | OpenBMB | 0.0 Agentic | 25.8 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 260 | Ministral 3 (14B Reasoning 2512) ministral-14b-latest multimodalvisionmulti-input reasoning | Mistral AI | 0.0 Agentic | 35.4 | 0.0 | 0.0 | 0.0 | 0.0 |
Llama 3.2 11B Instruct
Meta
0.0
N/A
Llama 3.2 3B Instruct
Meta
0.0
N/A
Llama 3.2 90B Instruct
Meta
0.0
N/A
Want benchmark charts, model comparison, and pricing analytics?
Sign in to access the full interactive leaderboard with deep benchmark breakdowns and model comparison tools.
Open full leaderboardRankings are based on multi-dimensional evaluation across benchmark quality, inference efficiency, and cost-per-output. Scores are updated continuously and may differ from individual third-party benchmarks.
| N/A |
| N/A |
| N/A |
Llama 3.3 70B Instruct
Meta
0.0
N/A
Llama-3.3 Nemotron Super 49B v1
NVIDIA
0.0
N/A
Llama 4 Maverick
Meta
0.0
N/A
Llama 4 Scout
Meta
0.0
N/A
LongCat-Flash-Thinking
Meituan
0.0
N/A
Magistral Medium
Mistral AI
0.0
N/A
Magistral Small 2506
Mistral AI
0.0
N/A
MAI-Code-1-Flash
Microsoft
0.0
N/A
MAI-Thinking-1
Microsoft
0.0
N/A
MedGemma 4B IT
0.0
N/A
Mercury 2
Inception
0.0
$0.25 in / $0.75 out
MiMo-V2.5
Xiaomi
0.0
$0.168 in / $0.336 out
MiMo-V2.5-Pro
Xiaomi
0.0
$0.435 in / $0.87 out
MiMo-V2-Omni
Xiaomi
0.0
N/A
MiMo-V2-Pro
Xiaomi
0.0
N/A
MiniCPM-SALA
OpenBMB
0.0
N/A
Ministral 3 (14B Reasoning 2512)
Mistral AI
0.0
N/A