Every major AI model ranked across benchmark quality, inference speed, agentic capability, programming aptitude, and cost efficiency — updated continuously from published evaluation data.
334
Tracked models
29
Providers
286
Benchmarked
13.1
Avg. index
334 models
| Rank | Model | Provider | Score | Benchmarks | Inference | Agentic | Programming | Value | Price |
|---|---|---|---|---|---|---|---|---|---|
| 221 | Llama 4 Scout llama-4-scout multimodalvisionmulti-input reasoning | Meta | 0.0 Inference | 27.6 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 222 | LongCat-Flash-Chat longcat-flash-chat codeprogrammingtool use | Meituan | 0.0 Inference | 26.0 | 0.0 | 48.1 | 36.6 | 0.0 | N/A |
| 223 | LongCat-Flash-Thinking longcat-flash-thinking codeprogrammingtool use | Meituan | 0.0 Inference | 48.2 | 0.0 | 0.0 | 18.4 | 0.0 | |
| 224 | LongCat-Flash-Thinking-2601 longcat-flash-thinking-2601 codeprogrammingtool use | Meituan | 0.0 Inference | 52.9 | 0.0 | 25.6 | 33.5 | 0.0 | |
| 225 | Magistral Medium magistral-medium multimodalvisionmulti-input reasoning | Mistral AI | 0.0 Inference | 20.6 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 226 | Magistral Small 2506 magistral-small-2506 textinference | Mistral AI | 0.0 Inference | 22.7 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 227 | MAI-Code-1-Flash mai-code-1-flash codeprogrammingtool use | Microsoft | 0.0 Inference | 32.2 | 0.0 | 0.0 | 21.4 | 0.0 | N/A |
| 228 | MAI-Thinking-1 mai-thinking-1 codeprogrammingtool use | Microsoft | 0.0 Inference | 60.1 | 0.0 | 0.0 | 32.2 | 0.0 | N/A |
| 229 | MedGemma 4B IT medgemma-4b-it multimodalvisionmulti-input reasoning | Google | 0.0 Inference | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 230 | MiMo-V2-Flash mimo-v2-flash codeprogrammingtool use | Xiaomi | 0.0 Inference | 50.6 | 0.0 | 23.5 | 35.8 | 0.0 | N/A |
| 231 | MiMo-V2-Omni mimo-v2-omni multimodalvisionmulti-input reasoning | Xiaomi | 0.0 Inference | 0.0 | 0.0 | 0.0 | 50.9 | 0.0 | N/A |
| 232 | MiMo-V2-Pro mimo-v2-pro codeprogrammingtool use | Xiaomi | 0.0 Inference | 0.0 | 0.0 | 0.0 | 61.9 | 0.0 | N/A |
| 233 | MiniCPM-SALA minicpm-sala textinference | OpenBMB | 0.0 Inference | 25.8 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 234 | MiniMax M1 40K minimax-m1-40k codeprogrammingtool use | MiniMax | 0.0 Inference | 21.3 | 0.0 | 26.8 | 15.5 | 0.0 | N/A |
| 235 | MiniMax M1 80K minimax-m1-80k codeprogrammingtool use | MiniMax | 0.0 Inference | 22.8 | 0.0 | 20.9 | 16.2 | 0.0 | N/A |
| 236 | Ministral 3 (14B Reasoning 2512) ministral-14b-latest multimodalvisionmulti-input reasoning | Mistral AI | 0.0 Inference | 35.4 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 237 | Ministral 3 (14B Base 2512) ministral-3-14b-base-2512 multimodalvisionmulti-input reasoning | Mistral AI | 0.0 Inference | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 238 | MiniStral 3 (14B Instruct 2512) ministral-3-14b-instruct-2512 multimodalvisionmulti-input reasoning | Mistral AI | 0.0 Inference | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 239 | Ministral 3 (3B Base 2512) ministral-3-3b-base-2512 multimodalvisionmulti-input reasoning | Mistral AI | 0.0 Inference | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | |
| 240 | Ministral 3 (3B Instruct 2512) ministral-3-3b-instruct-2512 multimodalvisionmulti-input reasoning | Mistral AI | 0.0 Inference | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 |
Llama 4 Scout
Meta
0.0
N/A
LongCat-Flash-Chat
Meituan
0.0
N/A
LongCat-Flash-Thinking
Meituan
0.0
N/A
Want benchmark charts, model comparison, and pricing analytics?
Sign in to access the full interactive leaderboard with deep benchmark breakdowns and model comparison tools.
Open full leaderboardRankings are based on multi-dimensional evaluation across benchmark quality, inference efficiency, and cost-per-output. Scores are updated continuously and may differ from individual third-party benchmarks.
| N/A |
| N/A |
| N/A |
| N/A |
| N/A |
| N/A |
| N/A |
| N/A |
LongCat-Flash-Thinking-2601
Meituan
0.0
N/A
Magistral Medium
Mistral AI
0.0
N/A
Magistral Small 2506
Mistral AI
0.0
N/A
MAI-Code-1-Flash
Microsoft
0.0
N/A
MAI-Thinking-1
Microsoft
0.0
N/A
MedGemma 4B IT
0.0
N/A
MiMo-V2-Flash
Xiaomi
0.0
N/A
MiMo-V2-Omni
Xiaomi
0.0
N/A
MiMo-V2-Pro
Xiaomi
0.0
N/A
MiniCPM-SALA
OpenBMB
0.0
N/A
MiniMax M1 40K
MiniMax
0.0
N/A
MiniMax M1 80K
MiniMax
0.0
N/A
Ministral 3 (14B Reasoning 2512)
Mistral AI
0.0
N/A
Ministral 3 (14B Base 2512)
Mistral AI
0.0
N/A
MiniStral 3 (14B Instruct 2512)
Mistral AI
0.0
N/A
Ministral 3 (3B Base 2512)
Mistral AI
0.0
N/A
Ministral 3 (3B Instruct 2512)
Mistral AI
0.0
N/A