Skytells
HomeModelsCLIChangelog
  • Home
  • Models
  • CLI
  • Changelog
Skytells

Addressing the world's greatest challenges with AI. Enterprise research, foundation models, and infrastructure trusted by organizations worldwide since 2012.

Get Started

  • Console
  • Learn
  • Documentation
  • API Reference
  • Pricing
  • ModelsNew

Platform

  • Cloud AgentsNew
  • AI Solutions
  • Infrastructure
  • Edge Network
  • Trust Center
  • CLI

Resources

  • Blog
  • Changelog
  • AI Leaderboard
  • Research
  • Status

Company

  • About
  • Careers
  • Legal
  • Privacy Policy

© 2012–2026 Skytells, Inc. All rights reserved.

Live rankings

AI Model Leaderboard

Every major AI model ranked across benchmark quality, inference speed, agentic capability, programming aptitude, and cost efficiency — updated continuously from published evaluation data.

Explore full leaderboardBrowse model catalog

334

Tracked models

29

Providers

286

Benchmarked

15.1

Avg. index

OverallBenchmarksInferenceAgenticProgrammingValue / Price

334 models

RankModelProviderScoreBenchmarksInferenceAgenticProgrammingValuePrice
121

DeepSeek-R1-0528

deepseek-r1-0528

codeprogrammingtool use
DeepSeek

5.7

Programming

47.90.00.05.70.0N/A
122

o1

o1-2024-12-17

multimodalvisionmulti-input reasoning
OpenAI

5.6

Programming

41.90.044.75.60.0N/A
123

GPT-4.5

gpt-4.5

multimodalvisionmulti-input reasoning
OpenAI

5.2

Programming

41.10.035.85.20.0N/A
124

Sarvam-30B

sarvam-30b

codeprogrammingtool use
SSarvam AI

4.4

Programming

44.90.06.44.40.0N/A
125

Nemotron 3 Nano (30B A3B)

nemotron-3-nano-30b-a3b

codeprogrammingtool use
NNVIDIA

3.8

Programming

43.632.93.03.8100.0$0.06 in / $0.24 out
126

GPT-4o

gpt-4o-2024-08-06

multimodalvisionmulti-input reasoning
OpenAI

3.7

Programming

29.139.614.93.731.2
127

Gemini 2.5 Flash-Lite

gemini-2.5-flash-lite

multimodalvisionmulti-input reasoning
Google

2.9

Programming

20.30.00.02.90.0
128

GPT-4.1 mini

gpt-4.1-mini-2025-04-14

multimodalvisionmulti-input reasoning
OpenAI

2.2

Programming

19.284.68.92.269.5
129

Gemini Diffusion

gemini-diffusion

codeprogrammingtool use
Google

1.5

Programming

6.50.00.01.50.0N/A
130

DeepSeek-V2.5

deepseek-v2.5

codeprogrammingtool use
DeepSeek

0.7

Programming

0.00.00.00.70.0N/A
131

ChatGPT-4o Latest

chatgpt-4o-latest

multimodalvisionmulti-input reasoning
OpenAI

0.0

Programming

53.00.00.00.00.0
132

Claude 3.5 Sonnet

claude-3-5-sonnet-20240620

multimodalvisionmulti-input reasoning
Anthropic

0.0

Programming

23.30.00.00.00.0
133

Claude 3 Haiku

claude-3-haiku-20240307

multimodalvisionmulti-input reasoning
Anthropic

0.0

Programming

5.30.00.00.00.0
134

Claude 3 Opus

claude-3-opus-20240229

multimodalvisionmulti-input reasoning
Anthropic

0.0

Programming

17.70.00.00.00.0
135

Claude 3 Sonnet

claude-3-sonnet-20240229

multimodalvisionmulti-input reasoning
Anthropic

0.0

Programming

9.20.00.00.00.0
136

Codestral-22B

codestral-22b

textinference
Mistral AI

0.0

Programming

0.00.00.00.00.0N/A
137

Command R+

command-r-plus-04-2024

textinference
Cohere

0.0

Programming

0.00.00.00.00.0N/A
138

DeepSeek-V3.2 (Non-thinking)

deepseek-chat

textinference
DeepSeek

0.0

Programming

0.052.00.00.082.7$0.28 in / $0.42 out
139

DeepSeek-R1

deepseek-r1

textinference
DeepSeek

0.0

Programming

0.00.00.00.00.0N/A
140

DeepSeek R1 Distill Llama 70B

deepseek-r1-distill-llama-70b

textinference
DeepSeek

0.0

Programming

26.40.00.00.00.0N/A
121

DeepSeek-R1-0528

DeepSeek

5.7

N/A

122

o1

OpenAI

5.6

N/A

123

GPT-4.5

OpenAI

5.2

N/A

124

Page 7 of 17 · 334 models

PreviousNext

Want benchmark charts, model comparison, and pricing analytics?

Sign in to access the full interactive leaderboard with deep benchmark breakdowns and model comparison tools.

Open full leaderboard

Rankings are based on multi-dimensional evaluation across benchmark quality, inference efficiency, and cost-per-output. Scores are updated continuously and may differ from individual third-party benchmarks.

$2.5 in / $10 out
N/A
$0.4 in / $1.6 out
N/A
N/A
N/A
N/A
N/A
S

Sarvam-30B

Sarvam AI

4.4

N/A

125
N

Nemotron 3 Nano (30B A3B)

NVIDIA

3.8

$0.06 in / $0.24 out

126

GPT-4o

OpenAI

3.7

$2.5 in / $10 out

127

Gemini 2.5 Flash-Lite

Google

2.9

N/A

128

GPT-4.1 mini

OpenAI

2.2

$0.4 in / $1.6 out

129

Gemini Diffusion

Google

1.5

N/A

130

DeepSeek-V2.5

DeepSeek

0.7

N/A

131

ChatGPT-4o Latest

OpenAI

0.0

N/A

132

Claude 3.5 Sonnet

Anthropic

0.0

N/A

133

Claude 3 Haiku

Anthropic

0.0

N/A

134

Claude 3 Opus

Anthropic

0.0

N/A

135

Claude 3 Sonnet

Anthropic

0.0

N/A

136

Codestral-22B

Mistral AI

0.0

N/A

137

Command R+

Cohere

0.0

N/A

138

DeepSeek-V3.2 (Non-thinking)

DeepSeek

0.0

$0.28 in / $0.42 out

139

DeepSeek-R1

DeepSeek

0.0

N/A

140

DeepSeek R1 Distill Llama 70B

DeepSeek

0.0

N/A