Every major AI model ranked across benchmark quality, inference speed, agentic capability, programming aptitude, and cost efficiency — updated continuously from published evaluation data.
334
Tracked models
29
Providers
286
Benchmarked
15.1
Avg. index
334 models
| Rank | Model | Provider | Score | Benchmarks | Inference | Agentic | Programming | Value | Price |
|---|---|---|---|---|---|---|---|---|---|
| 141 | DeepSeek R1 Distill Llama 8B deepseek-r1-distill-llama-8b textinference | DeepSeek | 0.0 Programming | 16.3 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 142 | DeepSeek R1 Distill Qwen 14B deepseek-r1-distill-qwen-14b textinference | DeepSeek | 0.0 Programming | 22.7 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 143 | DeepSeek R1 Distill Qwen 1.5B deepseek-r1-distill-qwen-1.5b textinference | DeepSeek | 0.0 Programming | 5.6 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 144 | DeepSeek R1 Distill Qwen 32B deepseek-r1-distill-qwen-32b textinference | DeepSeek | 0.0 Programming | 24.4 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 145 | DeepSeek R1 Distill Qwen 7B deepseek-r1-distill-qwen-7b textinference | DeepSeek | 0.0 Programming | 16.8 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 146 | DeepSeek R1 Zero deepseek-r1-zero textinference | DeepSeek | 0.0 Programming | 36.8 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 147 | DeepSeek-V3 0324 deepseek-v3-0324 textinference | DeepSeek | 0.0 Programming | 30.4 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 148 | DeepSeek VL2 deepseek-vl2 multimodalvisionmulti-input reasoning | DeepSeek | 0.0 Programming | 6.8 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 149 | DeepSeek VL2 Small deepseek-vl2-small multimodalvisionmulti-input reasoning | DeepSeek | 0.0 Programming | 4.6 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 150 | DeepSeek VL2 Tiny deepseek-vl2-tiny multimodalvisionmulti-input reasoning | DeepSeek | 0.0 Programming | 1.1 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 151 | DiffusionGemma 26B-A4B diffusiongemma-26b-a4b-it multimodalvisionmulti-input reasoning | Google | 0.0 Programming | 22.9 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 152 | ERNIE 4.5 ernie-4.5 textinference | Baidu | 0.0 Programming | 22.8 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 153 | ERNIE 5.0 ernie-5.0 multimodalvisionmulti-input reasoning | Baidu | 0.0 Programming | 56.8 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 154 | Gemini 1.0 Pro gemini-1.0-pro multimodalvisionmulti-input reasoning | Google | 0.0 Programming | 3.0 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 155 | Gemini 1.5 Flash gemini-1.5-flash multimodalvisionmulti-input reasoning | Google | 0.0 Programming | 21.8 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 156 | Gemini 1.5 Flash 8B gemini-1.5-flash-8b multimodalvisionmulti-input reasoning | Google | 0.0 Programming | 9.9 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 157 | Gemini 1.5 Pro gemini-1.5-pro multimodalvisionmulti-input reasoning | Google | 0.0 Programming | 26.2 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 158 | Gemini 2.0 Flash gemini-2.0-flash multimodalvisionmulti-input reasoning | Google | 0.0 Programming | 31.7 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 159 | Gemini 2.0 Flash-Lite gemini-2.0-flash-lite multimodalvisionmulti-input reasoning | Google | 0.0 Programming | 24.4 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 160 | Gemini 2.0 Flash Thinking gemini-2.0-flash-thinking multimodalvisionmulti-input reasoning | Google | 0.0 Programming | 44.9 | 0.0 | 0.0 | 0.0 | 0.0 |
DeepSeek R1 Distill Llama 8B
DeepSeek
0.0
N/A
DeepSeek R1 Distill Qwen 14B
DeepSeek
0.0
N/A
DeepSeek R1 Distill Qwen 1.5B
DeepSeek
0.0
N/A
Want benchmark charts, model comparison, and pricing analytics?
Sign in to access the full interactive leaderboard with deep benchmark breakdowns and model comparison tools.
Open full leaderboardRankings are based on multi-dimensional evaluation across benchmark quality, inference efficiency, and cost-per-output. Scores are updated continuously and may differ from individual third-party benchmarks.
| N/A |
| N/A |
| N/A |
| N/A |
| N/A |
| N/A |
| N/A |
| N/A |
| N/A |
| N/A |
DeepSeek R1 Distill Qwen 32B
DeepSeek
0.0
N/A
DeepSeek R1 Distill Qwen 7B
DeepSeek
0.0
N/A
DeepSeek R1 Zero
DeepSeek
0.0
N/A
DeepSeek-V3 0324
DeepSeek
0.0
N/A
DeepSeek VL2
DeepSeek
0.0
N/A
DeepSeek VL2 Small
DeepSeek
0.0
N/A
DeepSeek VL2 Tiny
DeepSeek
0.0
N/A
DiffusionGemma 26B-A4B
0.0
N/A
ERNIE 4.5
Baidu
0.0
N/A
ERNIE 5.0
Baidu
0.0
N/A
Gemini 1.0 Pro
0.0
N/A
Gemini 1.5 Flash
0.0
N/A
Gemini 1.5 Flash 8B
0.0
N/A
Gemini 1.5 Pro
0.0
N/A
Gemini 2.0 Flash
0.0
N/A
Gemini 2.0 Flash-Lite
0.0
N/A
Gemini 2.0 Flash Thinking
0.0
N/A