Every major AI model ranked across benchmark quality, inference speed, agentic capability, programming aptitude, and cost efficiency — updated continuously from published evaluation data.
334
Tracked models
29
Providers
286
Benchmarked
12.6
Avg. index
334 models
| Rank | Model | Provider | Score | Benchmarks | Inference | Agentic | Programming | Value | Price |
|---|---|---|---|---|---|---|---|---|---|
| 321 | Qwen3 VL 30B A3B Thinking qwen3-vl-30b-a3b-thinking multimodalvisionmulti-input reasoning | Alibaba Cloud / Qwen Team | 0.0 Value / Price | 32.7 | 0.0 | 19.2 | 0.0 | 0.0 | N/A |
| 322 | Qwen3 VL 32B Instruct qwen3-vl-32b-instruct multimodalvisionmulti-input reasoning | Alibaba Cloud / Qwen Team | 0.0 Value / Price | 26.5 | 0.0 | 25.1 | 0.0 | 0.0 | |
| 323 | Qwen3 VL 32B Thinking qwen3-vl-32b-thinking multimodalvisionmulti-input reasoning | Alibaba Cloud / Qwen Team | 0.0 Value / Price | 41.2 | 0.0 | 31.1 | 0.0 | 0.0 | |
| 324 | Qwen3 VL 8B Instruct qwen3-vl-8b-instruct multimodalvisionmulti-input reasoning | Alibaba Cloud / Qwen Team | 0.0 Value / Price | 8.0 | 0.0 | 24.0 | 0.0 | 0.0 | |
| 325 | Qwen3 VL 8B Thinking qwen3-vl-8b-thinking multimodalvisionmulti-input reasoning | Alibaba Cloud / Qwen Team | 0.0 Value / Price | 32.8 | 0.0 | 21.1 | 0.0 | 0.0 | |
| 326 | QwQ-32B qwq-32b textinference | Alibaba Cloud / Qwen Team | 0.0 Value / Price | 26.4 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 327 | QwQ-32B-Preview qwq-32b-preview textinference | Alibaba Cloud / Qwen Team | 0.0 Value / Price | 26.4 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 328 | Sarvam-105B sarvam-105b codeprogrammingtool use | Sarvam AI | 0.0 Value / Price | 41.1 | 0.0 | 16.7 | 10.3 | 0.0 | N/A |
| 329 | Sarvam-30B sarvam-30b codeprogrammingtool use | Sarvam AI | 0.0 Value / Price | 44.9 | 0.0 | 6.4 | 4.4 | 0.0 | N/A |
| 330 | Seed 2.0 Lite seed-2.0-lite multimodalvisionmulti-input reasoning | ByteDance | 0.0 Value / Price | 56.5 | 0.0 | 0.0 | 46.1 | 0.0 | N/A |
| 331 | Seed 2.1 Pro seed-2.1-pro multimodalvisionmulti-input reasoning | ByteDance | 0.0 Value / Price | 69.2 | 0.0 | 75.6 | 60.2 | 0.0 | N/A |
| 332 | Seed 2.1 Turbo seed-2.1-turbo multimodalvisionmulti-input reasoning | ByteDance | 0.0 Value / Price | 66.3 | 0.0 | 63.1 | 53.8 | 0.0 | N/A |
| 333 | Step3-VL-10B step3-vl-10b multimodalvisionmulti-input reasoning | StepFun | 0.0 Value / Price | 46.9 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
| 334 | U2 u2 textinference | Unisound | 0.0 Value / Price | 0.0 | 0.0 | 0.0 | 0.0 | 0.0 | N/A |
Qwen3 VL 30B A3B Thinking
Alibaba Cloud / Qwen Team
0.0
N/A
Qwen3 VL 32B Instruct
Alibaba Cloud / Qwen Team
0.0
N/A
Qwen3 VL 32B Thinking
Alibaba Cloud / Qwen Team
0.0
N/A
Page 17 of 17 · 334 models
Want benchmark charts, model comparison, and pricing analytics?
Sign in to access the full interactive leaderboard with deep benchmark breakdowns and model comparison tools.
Open full leaderboardRankings are based on multi-dimensional evaluation across benchmark quality, inference efficiency, and cost-per-output. Scores are updated continuously and may differ from individual third-party benchmarks.
| N/A |
| N/A |
| N/A |
| N/A |
Qwen3 VL 8B Instruct
Alibaba Cloud / Qwen Team
0.0
N/A
Qwen3 VL 8B Thinking
Alibaba Cloud / Qwen Team
0.0
N/A
QwQ-32B
Alibaba Cloud / Qwen Team
0.0
N/A
QwQ-32B-Preview
Alibaba Cloud / Qwen Team
0.0
N/A
Sarvam-105B
Sarvam AI
0.0
N/A
Sarvam-30B
Sarvam AI
0.0
N/A
Seed 2.0 Lite
ByteDance
0.0
N/A
Seed 2.1 Pro
ByteDance
0.0
N/A
Seed 2.1 Turbo
ByteDance
0.0
N/A
Step3-VL-10B
StepFun
0.0
N/A
U2
Unisound
0.0
N/A