Skytells
HomeModelsCLIChangelog
  • Home
  • Models
  • CLI
  • Changelog
Skytells

Addressing the world's greatest challenges with AI. Enterprise research, foundation models, and infrastructure trusted by organizations worldwide since 2012.

Get Started

  • Console
  • Learn
  • Documentation
  • API Reference
  • Pricing
  • ModelsNew

Platform

  • Cloud AgentsNew
  • AI Solutions
  • Infrastructure
  • Edge Network
  • Trust Center
  • CLI

Resources

  • Blog
  • Changelog
  • AI Leaderboard
  • Research
  • Status

Company

  • About
  • Careers
  • Legal
  • Privacy Policy

© 2012–2026 Skytells, Inc. All rights reserved.

Live rankings

AI Model Leaderboard

Every major AI model ranked across benchmark quality, inference speed, agentic capability, programming aptitude, and cost efficiency — updated continuously from published evaluation data.

Explore full leaderboardBrowse model catalog

334

Tracked models

29

Providers

286

Benchmarked

15.1

Avg. index

OverallBenchmarksInferenceAgenticProgrammingValue / Price

334 models

RankModelProviderScoreBenchmarksInferenceAgenticProgrammingValuePrice
141

DeepSeek R1 Distill Llama 8B

deepseek-r1-distill-llama-8b

textinference
DeepSeek

0.0

Programming

16.30.00.00.00.0N/A
142

DeepSeek R1 Distill Qwen 14B

deepseek-r1-distill-qwen-14b

textinference
DeepSeek

0.0

Programming

22.70.00.00.00.0N/A
143

DeepSeek R1 Distill Qwen 1.5B

deepseek-r1-distill-qwen-1.5b

textinference
DeepSeek

0.0

Programming

5.60.00.00.00.0N/A
144

DeepSeek R1 Distill Qwen 32B

deepseek-r1-distill-qwen-32b

textinference
DeepSeek

0.0

Programming

24.40.00.00.00.0N/A
145

DeepSeek R1 Distill Qwen 7B

deepseek-r1-distill-qwen-7b

textinference
DeepSeek

0.0

Programming

16.80.00.00.00.0N/A
146

DeepSeek R1 Zero

deepseek-r1-zero

textinference
DeepSeek

0.0

Programming

36.80.00.00.00.0N/A
147

DeepSeek-V3 0324

deepseek-v3-0324

textinference
DeepSeek

0.0

Programming

30.40.00.00.00.0N/A
148

DeepSeek VL2

deepseek-vl2

multimodalvisionmulti-input reasoning
DeepSeek

0.0

Programming

6.80.00.00.00.0N/A
149

DeepSeek VL2 Small

deepseek-vl2-small

multimodalvisionmulti-input reasoning
DeepSeek

0.0

Programming

4.60.00.00.00.0
150

DeepSeek VL2 Tiny

deepseek-vl2-tiny

multimodalvisionmulti-input reasoning
DeepSeek

0.0

Programming

1.10.00.00.00.0
151

DiffusionGemma 26B-A4B

diffusiongemma-26b-a4b-it

multimodalvisionmulti-input reasoning
Google

0.0

Programming

22.90.00.00.00.0
152

ERNIE 4.5

ernie-4.5

textinference
BBaidu

0.0

Programming

22.80.00.00.00.0N/A
153

ERNIE 5.0

ernie-5.0

multimodalvisionmulti-input reasoning
BBaidu

0.0

Programming

56.80.00.00.00.0N/A
154

Gemini 1.0 Pro

gemini-1.0-pro

multimodalvisionmulti-input reasoning
Google

0.0

Programming

3.00.00.00.00.0
155

Gemini 1.5 Flash

gemini-1.5-flash

multimodalvisionmulti-input reasoning
Google

0.0

Programming

21.80.00.00.00.0
156

Gemini 1.5 Flash 8B

gemini-1.5-flash-8b

multimodalvisionmulti-input reasoning
Google

0.0

Programming

9.90.00.00.00.0
157

Gemini 1.5 Pro

gemini-1.5-pro

multimodalvisionmulti-input reasoning
Google

0.0

Programming

26.20.00.00.00.0
158

Gemini 2.0 Flash

gemini-2.0-flash

multimodalvisionmulti-input reasoning
Google

0.0

Programming

31.70.00.00.00.0
159

Gemini 2.0 Flash-Lite

gemini-2.0-flash-lite

multimodalvisionmulti-input reasoning
Google

0.0

Programming

24.40.00.00.00.0
160

Gemini 2.0 Flash Thinking

gemini-2.0-flash-thinking

multimodalvisionmulti-input reasoning
Google

0.0

Programming

44.90.00.00.00.0
141

DeepSeek R1 Distill Llama 8B

DeepSeek

0.0

N/A

142

DeepSeek R1 Distill Qwen 14B

DeepSeek

0.0

N/A

143

DeepSeek R1 Distill Qwen 1.5B

DeepSeek

0.0

N/A

Page 8 of 17 · 334 models

PreviousNext

Want benchmark charts, model comparison, and pricing analytics?

Sign in to access the full interactive leaderboard with deep benchmark breakdowns and model comparison tools.

Open full leaderboard

Rankings are based on multi-dimensional evaluation across benchmark quality, inference efficiency, and cost-per-output. Scores are updated continuously and may differ from individual third-party benchmarks.

N/A
N/A
N/A
N/A
N/A
N/A
N/A
N/A
N/A
N/A
144

DeepSeek R1 Distill Qwen 32B

DeepSeek

0.0

N/A

145

DeepSeek R1 Distill Qwen 7B

DeepSeek

0.0

N/A

146

DeepSeek R1 Zero

DeepSeek

0.0

N/A

147

DeepSeek-V3 0324

DeepSeek

0.0

N/A

148

DeepSeek VL2

DeepSeek

0.0

N/A

149

DeepSeek VL2 Small

DeepSeek

0.0

N/A

150

DeepSeek VL2 Tiny

DeepSeek

0.0

N/A

151

DiffusionGemma 26B-A4B

Google

0.0

N/A

152
B

ERNIE 4.5

Baidu

0.0

N/A

153
B

ERNIE 5.0

Baidu

0.0

N/A

154

Gemini 1.0 Pro

Google

0.0

N/A

155

Gemini 1.5 Flash

Google

0.0

N/A

156

Gemini 1.5 Flash 8B

Google

0.0

N/A

157

Gemini 1.5 Pro

Google

0.0

N/A

158

Gemini 2.0 Flash

Google

0.0

N/A

159

Gemini 2.0 Flash-Lite

Google

0.0

N/A

160

Gemini 2.0 Flash Thinking

Google

0.0

N/A