Skytells
  • Home
  • Changelog
Skytells

Addressing the world's greatest challenges with AI. Enterprise research, foundation models, and infrastructure trusted by organizations worldwide since 2012.

Get Started

  • Console
  • Learn
  • Pricing
  • ModelsNew
  • EveNew
  • CLI

Platform

  • DropsNew
  • Cognition
  • Cloud AgentsNew
  • AI Solutions
  • Infrastructure
  • Edge Network

Resources

  • Documentation
  • Blog
  • Changelog
  • AI Leaderboard
  • Research
  • Status

Company

  • About
  • Careers
  • Trust Center
  • Legal
  • Privacy Policy

© 2012–2026 Skytells, Inc. All rights reserved.

Live rankings

AI Model Leaderboard

Every major AI model ranked across benchmark quality, inference speed, agentic capability, programming aptitude, and cost efficiency — updated continuously from published evaluation data.

Explore full leaderboardBrowse model catalog

373

Tracked models

34

Providers

310

Benchmarked

30.2

Avg. index

OverallBenchmarksInferenceAgenticProgrammingValue / Price

373 models

RankModelProviderScoreBenchmarksInferenceAgenticProgrammingValuePrice
21

Muse Spark 1.1

muse-spark-1.1

multimodalvisionmulti-input reasoning
MMeta

65.8

overall

66.481.975.655.138.8$1.25 in / $4.25 out
22

Seed 2.1 Pro

seed-2.1-pro

multimodalvisionmulti-input reasoning
BByteDance

65.4

overall

67.20.068.459.90.0N/A
23

GLM-5.3

glm-5.3

textinference
ZZhipu AI

64.7

overall

68.881.959.90.037.4$1.4 in / $4.4 out
24

Qwen3.8 Flash

qwen3.8-flash

multimodalvisionmulti-input reasoning
AAlibaba Cloud / Qwen Team

63.5

overall

61.158.664.560.083.5$0.15 in / $0.47 out
25

Claude Fable 5

claude-fable-5

multimodalvisionmulti-input reasoning
Anthropic

62.5

overall

69.558.60.084.21.9
26

Qwen3.8-Flash-Next

qwen3.8-flash-next

multimodalvisionmulti-input reasoning
AAlibaba Cloud / Qwen Team

61.9

overall

61.10.064.560.00.0N/A
27

MiMo-V2-Pro

mimo-v2-pro

codeprogrammingtool use
Xiaomi

61.7

overall

0.00.00.061.70.0N/A
28

GPT-5.1 Codex High

gpt-5.1-codex-high

multimodalvisionmulti-input reasoning
OpenAI

61.4

overall

61.40.00.00.00.0
29

GPT-5 High

gpt-5-high-2025-08-07

multimodalvisionmulti-input reasoning
OpenAI

60.8

overall

60.80.00.00.00.0
30

GPT-5.5

gpt-5.5

multimodalvisionmulti-input reasoning
OpenAI

60.6

overall

74.295.254.949.25.8$5 in / $30 out
31

Claude Opus 4.8

claude-opus-4-8

multimodalvisionmulti-input reasoning
Anthropic

60.0

overall

72.526.768.781.89.5
32

Claude Opus 5

claude-opus-5

multimodalvisionmulti-input reasoning
Anthropic

59.9

overall

70.558.669.40.09.2
33

DeepSeek-V3.2 (Non-thinking)

deepseek-chat

textinference
DeepSeek

59.7

overall

0.048.70.00.077.3$0.28 in / $0.42 out
34

Qwen3.7 Max

qwen3.7-max

multimodalvisionmulti-input reasoning
AAlibaba Cloud / Qwen Team

59.3

overall

64.058.648.074.440.8$1.25 in / $3.75 out
35

Seed 2.1 Turbo

seed-2.1-turbo

multimodalvisionmulti-input reasoning
BByteDance

59.1

overall

63.60.057.655.10.0N/A
36

Kimi K2-Thinking-0905

kimi-k2-thinking-0905

codeprogrammingtool use
Moonshot AI

59.0

overall

64.90.051.259.90.0
37

GPT-5.5 Pro

gpt-5.5-pro

multimodalvisionmulti-input reasoning
OpenAI

58.9

overall

59.90.067.148.80.0N/A
38

GLM-5.2

glm-5.2

codeprogrammingtool use
ZZhipu AI

58.7

overall

66.881.939.157.947.6$0.95 in / $3 out
39

Gemini 3.8 Flash

gemini-3.8-flash

multimodalvisionmulti-input reasoning
Google

58.5

overall

50.681.90.00.043.2
40

Gemini 3 Pro

gemini-3-pro-preview

multimodalvisionmulti-input reasoning
Google

58.3

overall

68.80.051.052.90.0
21
M

Muse Spark 1.1

Meta

65.8

$1.25 in / $4.25 out

22
B

Seed 2.1 Pro

ByteDance

65.4

N/A

23
Z

GLM-5.3

Zhipu AI

64.7

$1.4 in / $4.4 out

24

Page 2 of 19 · 373 models

PreviousNext

Want benchmark charts, model comparison, and pricing analytics?

Sign in to access the full interactive leaderboard with deep benchmark breakdowns and model comparison tools.

Open full leaderboard

Rankings are based on multi-dimensional evaluation across benchmark quality, inference efficiency, and cost-per-output. Scores are updated continuously and may differ from individual third-party benchmarks.

$10 in / $50 out
N/A
N/A
$5 in / $25 out
$5 in / $25 out
N/A
$0.75 in / $3.75 out
N/A
A

Qwen3.8 Flash

Alibaba Cloud / Qwen Team

63.5

$0.15 in / $0.47 out

25

Claude Fable 5

Anthropic

62.5

$10 in / $50 out

26
A

Qwen3.8-Flash-Next

Alibaba Cloud / Qwen Team

61.9

N/A

27

MiMo-V2-Pro

Xiaomi

61.7

N/A

28

GPT-5.1 Codex High

OpenAI

61.4

N/A

29

GPT-5 High

OpenAI

60.8

N/A

30

GPT-5.5

OpenAI

60.6

$5 in / $30 out

31

Claude Opus 4.8

Anthropic

60.0

$5 in / $25 out

32

Claude Opus 5

Anthropic

59.9

$5 in / $25 out

33

DeepSeek-V3.2 (Non-thinking)

DeepSeek

59.7

$0.28 in / $0.42 out

34
A

Qwen3.7 Max

Alibaba Cloud / Qwen Team

59.3

$1.25 in / $3.75 out

35
B

Seed 2.1 Turbo

ByteDance

59.1

N/A

36

Kimi K2-Thinking-0905

Moonshot AI

59.0

N/A

37

GPT-5.5 Pro

OpenAI

58.9

N/A

38
Z

GLM-5.2

Zhipu AI

58.7

$0.95 in / $3 out

39

Gemini 3.8 Flash

Google

58.5

$0.75 in / $3.75 out

40

Gemini 3 Pro

Google

58.3

N/A