Skytells
HomeModelsCLIChangelog
  • Home
  • Models
  • CLI
  • Changelog
Skytells

Addressing the world's greatest challenges with AI. Enterprise research, foundation models, and infrastructure trusted by organizations worldwide since 2012.

Get Started

  • Console
  • Learn
  • Documentation
  • API Reference
  • Pricing
  • ModelsNew

Platform

  • Cloud AgentsNew
  • AI Solutions
  • Infrastructure
  • Edge Network
  • Trust Center
  • CLI

Resources

  • Blog
  • Changelog
  • AI Leaderboard
  • Research
  • Status

Company

  • About
  • Careers
  • Legal
  • Privacy Policy

© 2012–2026 Skytells, Inc. All rights reserved.

Live rankings

AI Model Leaderboard

Every major AI model ranked across benchmark quality, inference speed, agentic capability, programming aptitude, and cost efficiency — updated continuously from published evaluation data.

Explore full leaderboardBrowse model catalog

334

Tracked models

29

Providers

286

Benchmarked

12.2

Avg. index

OverallBenchmarksInferenceAgenticProgrammingValue / Price

334 models

RankModelProviderScoreBenchmarksInferenceAgenticProgrammingValuePrice
101

Nemotron 3 Ultra (550B A55B)

nemotron-3-ultra-550b-a55b

codeprogrammingtool use
NNVIDIA

11.5

Agentic

54.70.011.540.60.0N/A
102

Qwen3.6-35B-A3B

qwen3.6-35b-a3b

multimodalvisionmulti-input reasoning
AAlibaba Cloud / Qwen Team

9.8

Agentic

51.10.09.825.20.0N/A
103

GLM-4.7-Flash

glm-4.7-flash

codeprogrammingtool use
ZZhipu AI

9.0

Agentic

36.40.09.017.70.0N/A
104

GPT-4.1 mini

gpt-4.1-mini-2025-04-14

multimodalvisionmulti-input reasoning
OpenAI

8.9

Agentic

19.284.68.92.269.5
105

Nemotron 3 Super (120B A12B)

nemotron-3-super-120b-a12b

codeprogrammingtool use
NNVIDIA

7.6

Agentic

45.50.07.622.00.0N/A
106

GPT-5.4 nano

gpt-5.4-nano

multimodalvisionmulti-input reasoning
OpenAI

6.9

Agentic

41.844.56.98.276.8$0.2 in / $1.25 out
107

Sarvam-30B

sarvam-30b

codeprogrammingtool use
SSarvam AI

6.4

Agentic

44.90.06.44.40.0N/A
108

GPT OSS 20B

gpt-oss-20b

textinference
OpenAI

6.0

Agentic

23.70.06.00.00.0N/A
109

Kimi K2-Instruct-0905

kimi-k2-instruct-0905

codeprogrammingtool use
Moonshot AI

6.0

Agentic

23.20.06.017.10.0
110

Qwen2.5 VL 72B Instruct

qwen2.5-vl-72b

multimodalvisionmulti-input reasoning
AAlibaba Cloud / Qwen Team

5.2

Agentic

22.70.05.20.00.0N/A
111

DeepSeek-V3.2-Speciale

deepseek-v3.2-speciale

codeprogrammingtool use
DeepSeek

5.0

Agentic

50.70.05.042.00.0
112

Claude 3.5 Haiku

claude-3-5-haiku-20241022

codeprogrammingtool use
Anthropic

3.0

Agentic

9.90.03.06.60.0
113

Nemotron 3 Nano (30B A3B)

nemotron-3-nano-30b-a3b

codeprogrammingtool use
NNVIDIA

3.0

Agentic

43.632.93.03.8100.0$0.06 in / $0.24 out
114

Qwen2.5 VL 32B Instruct

qwen2.5-vl-32b

multimodalvisionmulti-input reasoning
AAlibaba Cloud / Qwen Team

1.5

Agentic

19.40.01.50.00.0N/A
115

ChatGPT-4o Latest

chatgpt-4o-latest

multimodalvisionmulti-input reasoning
OpenAI

0.0

Agentic

53.00.00.00.00.0
116

Claude 3.5 Sonnet

claude-3-5-sonnet-20240620

multimodalvisionmulti-input reasoning
Anthropic

0.0

Agentic

23.30.00.00.00.0
117

Claude 3 Haiku

claude-3-haiku-20240307

multimodalvisionmulti-input reasoning
Anthropic

0.0

Agentic

5.30.00.00.00.0
118

Claude 3 Opus

claude-3-opus-20240229

multimodalvisionmulti-input reasoning
Anthropic

0.0

Agentic

17.70.00.00.00.0
119

Claude 3 Sonnet

claude-3-sonnet-20240229

multimodalvisionmulti-input reasoning
Anthropic

0.0

Agentic

9.20.00.00.00.0
120

Claude Fable 5

claude-fable-5

multimodalvisionmulti-input reasoning
Anthropic

0.0

Agentic

70.862.80.084.20.0
101
N

Nemotron 3 Ultra (550B A55B)

NVIDIA

11.5

N/A

102
A

Qwen3.6-35B-A3B

Alibaba Cloud / Qwen Team

9.8

N/A

103
Z

GLM-4.7-Flash

Zhipu AI

9.0

N/A

104

Page 6 of 17 · 334 models

PreviousNext

Want benchmark charts, model comparison, and pricing analytics?

Sign in to access the full interactive leaderboard with deep benchmark breakdowns and model comparison tools.

Open full leaderboard

Rankings are based on multi-dimensional evaluation across benchmark quality, inference efficiency, and cost-per-output. Scores are updated continuously and may differ from individual third-party benchmarks.

$0.4 in / $1.6 out
N/A
N/A
N/A
N/A
N/A
N/A
N/A
N/A
$10 in / $50 out

GPT-4.1 mini

OpenAI

8.9

$0.4 in / $1.6 out

105
N

Nemotron 3 Super (120B A12B)

NVIDIA

7.6

N/A

106

GPT-5.4 nano

OpenAI

6.9

$0.2 in / $1.25 out

107
S

Sarvam-30B

Sarvam AI

6.4

N/A

108

GPT OSS 20B

OpenAI

6.0

N/A

109

Kimi K2-Instruct-0905

Moonshot AI

6.0

N/A

110
A

Qwen2.5 VL 72B Instruct

Alibaba Cloud / Qwen Team

5.2

N/A

111

DeepSeek-V3.2-Speciale

DeepSeek

5.0

N/A

112

Claude 3.5 Haiku

Anthropic

3.0

N/A

113
N

Nemotron 3 Nano (30B A3B)

NVIDIA

3.0

$0.06 in / $0.24 out

114
A

Qwen2.5 VL 32B Instruct

Alibaba Cloud / Qwen Team

1.5

N/A

115

ChatGPT-4o Latest

OpenAI

0.0

N/A

116

Claude 3.5 Sonnet

Anthropic

0.0

N/A

117

Claude 3 Haiku

Anthropic

0.0

N/A

118

Claude 3 Opus

Anthropic

0.0

N/A

119

Claude 3 Sonnet

Anthropic

0.0

N/A

120

Claude Fable 5

Anthropic

0.0

$10 in / $50 out