Every major AI model ranked across benchmark quality, inference speed, agentic capability, programming aptitude, and cost efficiency — updated continuously from published evaluation data.
334
Tracked models
29
Providers
286
Benchmarked
29.3
Avg. index
334 models
| Rank | Model | Provider | Score | Benchmarks | Inference | Agentic | Programming | Value | Price |
|---|---|---|---|---|---|---|---|---|---|
| 281 | Jamba 1.5 Mini jamba-1.5-mini textinference | AI21 Labs | 4.3 overall | 4.3 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 282 | GPT-5.1 Codex Mini gpt-5.1-codex-mini multimodalvisionmulti-input reasoning | OpenAI | 3.8 overall | 3.8 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 283 | Llama 3.2 11B Instruct llama-3.2-11b-instruct multimodalvisionmulti-input reasoning | Meta | 3.8 overall | 3.8 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 284 | Gemini 1.0 Pro gemini-1.0-pro multimodalvisionmulti-input reasoning | Google | 3.0 overall | 3.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 285 | Llama 3.1 8B Instruct llama-3.1-8b-instruct textinference | Meta | 3.0 overall | 3.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 286 | Phi-3.5-mini-instruct phi-3.5-mini-instruct multimodalvisionmulti-input reasoning | Microsoft | 2.4 overall | 2.4 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 287 | Phi-3.5-vision-instruct phi-3.5-vision-instruct multimodalvisionmulti-input reasoning | Microsoft | 2.3 overall | 2.3 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 288 | Qwen2 7B Instruct qwen2-7b-instruct textinference | Alibaba Cloud / Qwen Team | 2.2 overall | 2.2 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 289 | Phi 4 Mini phi-4-mini textinference | Microsoft | 1.9 overall | 1.9 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 290 | Gemma 3n E4B Instructed gemma-3n-e4b-it multimodalvisionmulti-input reasoning | Google | 1.2 overall | 1.2 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 291 | Gemma 3n E4B Instructed LiteRT Preview gemma-3n-e4b-it-litert-preview multimodalvisionmulti-input reasoning | Google | 1.2 overall | 1.2 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 292 | DeepSeek VL2 Tiny deepseek-vl2-tiny multimodalvisionmulti-input reasoning | DeepSeek | 1.1 overall | 1.1 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 293 | Gemma 3n E2B Instructed gemma-3n-e2b-it multimodalvisionmulti-input reasoning | Google | 1.0 overall | 1.0 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 294 | Gemma 3n E2B Instructed LiteRT (Preview) gemma-3n-e2b-it-litert-preview multimodalvisionmulti-input reasoning | Google | 1.0 overall | 1.0 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 295 | Gemma 3 1B gemma-3-1b-it textinference | Google | 0.9 overall | 0.9 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 296 | DeepSeek-V2.5 deepseek-v2.5 codeprogrammingtool use | DeepSeek | 0.7 overall | 0.0 | 0.0 | 0.0 | 0.7 | 0.0 | N/A |
| 297 | Codestral-22B codestral-22b textinference | Mistral AI | 0.0 overall | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 298 | Command R+ command-r-plus-04-2024 textinference | Cohere | 0.0 overall | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 299 | DeepSeek-R1 deepseek-r1 textinference | DeepSeek | 0.0 overall | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 300 | Gemini 3.5 Flash Cyber gemini-3.5-flash-cyber textinference | Google | 0.0 overall | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
Jamba 1.5 Mini
AI21 Labs
4.3
N/A
GPT-5.1 Codex Mini
OpenAI
3.8
N/A
Llama 3.2 11B Instruct
Meta
3.8
N/A
Want benchmark charts, model comparison, and pricing analytics?
Sign in to access the full interactive leaderboard with deep benchmark breakdowns and model comparison tools.
Open full leaderboardRankings are based on multi-dimensional evaluation across benchmark quality, inference efficiency, and cost-per-output. Scores are updated continuously and may differ from individual third-party benchmarks.
| N/A |
| N/A |
| N/A |
| N/A |
| N/A |
| N/A |
Gemini 1.0 Pro
3.0
N/A
Llama 3.1 8B Instruct
Meta
3.0
N/A
Phi-3.5-mini-instruct
Microsoft
2.4
N/A
Phi-3.5-vision-instruct
Microsoft
2.3
N/A
Qwen2 7B Instruct
Alibaba Cloud / Qwen Team
2.2
N/A
Phi 4 Mini
Microsoft
1.9
N/A
Gemma 3n E4B Instructed
1.2
N/A
Gemma 3n E4B Instructed LiteRT Preview
1.2
N/A
DeepSeek VL2 Tiny
DeepSeek
1.1
N/A
Gemma 3n E2B Instructed
1.0
N/A
Gemma 3n E2B Instructed LiteRT (Preview)
1.0
N/A
Gemma 3 1B
0.9
N/A
DeepSeek-V2.5
DeepSeek
0.7
N/A
Codestral-22B
Mistral AI
0.0
N/A
Command R+
Cohere
0.0
N/A
DeepSeek-R1
DeepSeek
0.0
N/A
Gemini 3.5 Flash Cyber
0.0
N/A