Skytells
HomeModelsCLIChangelog
  • Home
  • Models
  • CLI
  • Changelog
Skytells

Addressing the world's greatest challenges with AI. Enterprise research, foundation models, and infrastructure trusted by organizations worldwide since 2012.

Get Started

  • Console
  • Learn
  • Documentation
  • API Reference
  • Pricing
  • ModelsNew

Platform

  • Cloud AgentsNew
  • AI Solutions
  • Infrastructure
  • Edge Network
  • Trust Center
  • CLI

Resources

  • Blog
  • Changelog
  • AI Leaderboard
  • Research
  • Status

Company

  • About
  • Careers
  • Legal
  • Privacy Policy

© 2012–2026 Skytells, Inc. All rights reserved.

Live rankings

AI Model Leaderboard

Every major AI model ranked across benchmark quality, inference speed, agentic capability, programming aptitude, and cost efficiency — updated continuously from published evaluation data.

Explore full leaderboardBrowse model catalog

334

Tracked models

29

Providers

286

Benchmarked

13.1

Avg. index

OverallBenchmarksInferenceAgenticProgrammingValue / Price

334 models

RankModelProviderScoreBenchmarksInferenceAgenticProgrammingValuePrice
81

Claude Sonnet 4.6

claude-sonnet-4-6

multimodalvisionmulti-input reasoning
Anthropic

12.6

Inference

62.312.641.065.612.0$3 in / $15 out
82

GLM-5

glm-5

codeprogrammingtool use
ZZhipu AI

6.9

Inference

0.06.935.160.440.0$1 in / $3.2 out
83

Qwen3 32B

qwen3-32b

textinference
AAlibaba Cloud / Qwen Team

2.2

Inference

20.12.20.00.078.0$0.1 in / $0.3 out
84

ChatGPT-4o Latest

chatgpt-4o-latest

multimodalvisionmulti-input reasoning
OpenAI

0.0

Inference

53.00.00.00.00.0
85

Claude 3.5 Haiku

claude-3-5-haiku-20241022

codeprogrammingtool use
Anthropic

0.0

Inference

9.90.03.06.60.0
86

Claude 3.5 Sonnet

claude-3-5-sonnet-20240620

multimodalvisionmulti-input reasoning
Anthropic

0.0

Inference

23.30.00.00.00.0
87

Claude 3.5 Sonnet

claude-3-5-sonnet-20241022

multimodalvisionmulti-input reasoning
Anthropic

0.0

Inference

32.00.038.711.10.0
88

Claude 3.7 Sonnet

claude-3-7-sonnet-20250219

multimodalvisionmulti-input reasoning
Anthropic

0.0

Inference

42.30.049.137.70.0
89

Claude 3 Haiku

claude-3-haiku-20240307

multimodalvisionmulti-input reasoning
Anthropic

0.0

Inference

5.30.00.00.00.0
90

Claude 3 Opus

claude-3-opus-20240229

multimodalvisionmulti-input reasoning
Anthropic

0.0

Inference

17.70.00.00.00.0
91

Claude 3 Sonnet

claude-3-sonnet-20240229

multimodalvisionmulti-input reasoning
Anthropic

0.0

Inference

9.20.00.00.00.0
92

Claude Mythos Preview

claude-mythos-preview

multimodalvisionmulti-input reasoning
Anthropic

0.0

Inference

80.00.066.682.70.0
93

Claude Opus 4.1

claude-opus-4-1-20250805

multimodalvisionmulti-input reasoning
Anthropic

0.0

Inference

46.00.067.460.80.0
94

Claude Opus 4

claude-opus-4-20250514

multimodalvisionmulti-input reasoning
Anthropic

0.0

Inference

35.80.057.447.00.0
95

Claude Opus 4.5

claude-opus-4-5-20251101

multimodalvisionmulti-input reasoning
Anthropic

0.0

Inference

54.50.035.272.20.0
96

Claude Sonnet 4

claude-sonnet-4-20250514

multimodalvisionmulti-input reasoning
Anthropic

0.0

Inference

39.30.049.442.80.0
97

Codestral-22B

codestral-22b

textinference
Mistral AI

0.0

Inference

0.00.00.00.00.0N/A
98

Command A+

command-a-plus-05-2026

multimodalvisionmulti-input reasoning
Cohere

0.0

Inference

35.20.00.015.30.0
99

Command R+

command-r-plus-04-2024

textinference
Cohere

0.0

Inference

0.00.00.00.00.0N/A
100

DeepSeek-R1

deepseek-r1

textinference
DeepSeek

0.0

Inference

0.00.00.00.00.0N/A
81

Claude Sonnet 4.6

Anthropic

12.6

$3 in / $15 out

82
Z

GLM-5

Zhipu AI

6.9

$1 in / $3.2 out

83
A

Qwen3 32B

Alibaba Cloud / Qwen Team

2.2

$0.1 in / $0.3 out

84

Page 5 of 17 · 334 models

PreviousNext

Want benchmark charts, model comparison, and pricing analytics?

Sign in to access the full interactive leaderboard with deep benchmark breakdowns and model comparison tools.

Open full leaderboard

Rankings are based on multi-dimensional evaluation across benchmark quality, inference efficiency, and cost-per-output. Scores are updated continuously and may differ from individual third-party benchmarks.

N/A
N/A
N/A
N/A
N/A
N/A
N/A
N/A
N/A
N/A
N/A
N/A
N/A
N/A

ChatGPT-4o Latest

OpenAI

0.0

N/A

85

Claude 3.5 Haiku

Anthropic

0.0

N/A

86

Claude 3.5 Sonnet

Anthropic

0.0

N/A

87

Claude 3.5 Sonnet

Anthropic

0.0

N/A

88

Claude 3.7 Sonnet

Anthropic

0.0

N/A

89

Claude 3 Haiku

Anthropic

0.0

N/A

90

Claude 3 Opus

Anthropic

0.0

N/A

91

Claude 3 Sonnet

Anthropic

0.0

N/A

92

Claude Mythos Preview

Anthropic

0.0

N/A

93

Claude Opus 4.1

Anthropic

0.0

N/A

94

Claude Opus 4

Anthropic

0.0

N/A

95

Claude Opus 4.5

Anthropic

0.0

N/A

96

Claude Sonnet 4

Anthropic

0.0

N/A

97

Codestral-22B

Mistral AI

0.0

N/A

98

Command A+

Cohere

0.0

N/A

99

Command R+

Cohere

0.0

N/A

100

DeepSeek-R1

DeepSeek

0.0

N/A