Skytells
  • Home
  • Changelog
Skytells

Addressing the world's greatest challenges with AI. Enterprise research, foundation models, and infrastructure trusted by organizations worldwide since 2012.

Get Started

  • Console
  • Learn
  • Pricing
  • ModelsNew
  • EveNew
  • CLI

Platform

  • DropsNew
  • Cognition
  • Cloud AgentsNew
  • AI Solutions
  • Infrastructure
  • Edge Network

Resources

  • Documentation
  • Blog
  • Changelog
  • AI Leaderboard
  • Research
  • Status

Company

  • About
  • Careers
  • Trust Center
  • Legal
  • Privacy Policy

© 2012–2026 Skytells, Inc. All rights reserved.

Live rankings

AI Model Leaderboard

Every major AI model ranked across benchmark quality, inference speed, agentic capability, programming aptitude, and cost efficiency — updated continuously from published evaluation data.

Explore full leaderboardBrowse model catalog

373

Tracked models

34

Providers

310

Benchmarked

28.2

Avg. index

OverallBenchmarksInferenceAgenticProgrammingValue / Price

373 models

RankModelProviderScoreBenchmarksInferenceAgenticProgrammingValuePrice
1

Claude Mythos Preview

claude-mythos-preview

multimodalvisionmulti-input reasoning
Anthropic

79.5

Benchmarks

79.50.064.883.00.0N/A
2

GPT-5.6 Sol

gpt-5.6-sol

multimodalvisionmulti-input reasoning
OpenAI

77.6

Benchmarks

77.695.263.172.75.8
3

Kimi K3

kimi-k3

multimodalvisionmulti-input reasoning
Moonshot AI

75.3

Benchmarks

75.381.978.20.013.1
4

GPT-6 Astra

gpt-6-astra

multimodalvisionmulti-input reasoning
OpenAI

74.8

Benchmarks

74.895.274.60.01.9$10 in / $50 out
5

Claude Opus 4.7

claude-opus-4-7

multimodalvisionmulti-input reasoning
Anthropic

74.3

Benchmarks

74.326.752.977.59.5
6

GPT-5.5

gpt-5.5

multimodalvisionmulti-input reasoning
OpenAI

74.2

Benchmarks

74.295.254.949.25.8$5 in / $30 out
7

Claude Opus 4.8

claude-opus-4-8

multimodalvisionmulti-input reasoning
Anthropic

72.5

Benchmarks

72.526.768.781.89.5
8

GPT-5.6 Terra

gpt-5.6-terra

multimodalvisionmulti-input reasoning
OpenAI

72.5

Benchmarks

72.595.254.669.321.4
9

Claude Opus 4.6

claude-opus-4-6

multimodalvisionmulti-input reasoning
Anthropic

72.2

Benchmarks

72.226.748.871.99.5
10

Claude Fable 5.1

claude-fable-5-1

multimodalvisionmulti-input reasoning
Anthropic

71.6

Benchmarks

71.658.60.00.01.9
11

Claude Opus 5

claude-opus-5

multimodalvisionmulti-input reasoning
Anthropic

70.5

Benchmarks

70.558.669.40.09.2
12

Gemini 3.1 Pro

gemini-3.1-pro-preview

multimodalvisionmulti-input reasoning
Google

70.5

Benchmarks

70.557.546.262.921.9
13

Muse Spark 1.3

muse-spark-1.3

multimodalvisionmulti-input reasoning
MMeta

70.4

Benchmarks

70.481.90.00.095.1$0.1 in / $0.2 out
14

Qwen3.8 Max

qwen3.8-max

multimodalvisionmulti-input reasoning
AAlibaba Cloud / Qwen Team

70.3

Benchmarks

70.381.962.772.035.0$1.65 in / $4.951 out
15

GPT-5.2

gpt-5.2-2025-12-11

multimodalvisionmulti-input reasoning
OpenAI

69.8

Benchmarks

69.858.934.366.328.7
16

Claude Fable 5

claude-fable-5

multimodalvisionmulti-input reasoning
Anthropic

69.5

Benchmarks

69.558.60.084.21.9
17

GPT-5.4

gpt-5.4

texttext-to-textlanguage
OpenAI

69.3

Benchmarks

69.335.443.649.417.6
18

Grok 4.5

grok-4.5

multimodalvisionmulti-input reasoning
xAI

69.1

Benchmarks

69.137.90.069.233.5$2 in / $6 out
19

Gemini 3 Pro

gemini-3-pro-preview

multimodalvisionmulti-input reasoning
Google

68.8

Benchmarks

68.80.051.052.90.0
20

GLM-5.3

glm-5.3

textinference
ZZhipu AI

68.8

Benchmarks

68.881.959.90.037.4$1.4 in / $4.4 out
1

Claude Mythos Preview

Anthropic

79.5

N/A

2

GPT-5.6 Sol

OpenAI

77.6

$5 in / $30 out

3

Kimi K3

Moonshot AI

75.3

$3 in / $15 out

Page 1 of 19 · 373 models

Next

Want benchmark charts, model comparison, and pricing analytics?

Sign in to access the full interactive leaderboard with deep benchmark breakdowns and model comparison tools.

Open full leaderboard

Rankings are based on multi-dimensional evaluation across benchmark quality, inference efficiency, and cost-per-output. Scores are updated continuously and may differ from individual third-party benchmarks.

$5 in / $30 out
$3 in / $15 out
$5 in / $25 out
$5 in / $25 out
$2 in / $12 out
$5 in / $25 out
$10 in / $50 out
$5 in / $25 out
$2.5 in / $15 out
$1.75 in / $14 out
$10 in / $50 out
$2.5 in / $15 out
N/A
4

GPT-6 Astra

OpenAI

74.8

$10 in / $50 out

5

Claude Opus 4.7

Anthropic

74.3

$5 in / $25 out

6

GPT-5.5

OpenAI

74.2

$5 in / $30 out

7

Claude Opus 4.8

Anthropic

72.5

$5 in / $25 out

8

GPT-5.6 Terra

OpenAI

72.5

$2 in / $12 out

9

Claude Opus 4.6

Anthropic

72.2

$5 in / $25 out

10

Claude Fable 5.1

Anthropic

71.6

$10 in / $50 out

11

Claude Opus 5

Anthropic

70.5

$5 in / $25 out

12

Gemini 3.1 Pro

Google

70.5

$2.5 in / $15 out

13
M

Muse Spark 1.3

Meta

70.4

$0.1 in / $0.2 out

14
A

Qwen3.8 Max

Alibaba Cloud / Qwen Team

70.3

$1.65 in / $4.951 out

15

GPT-5.2

OpenAI

69.8

$1.75 in / $14 out

16

Claude Fable 5

Anthropic

69.5

$10 in / $50 out

17

GPT-5.4

OpenAI

69.3

$2.5 in / $15 out

18

Grok 4.5

xAI

69.1

$2 in / $6 out

19

Gemini 3 Pro

Google

68.8

N/A

20
Z

GLM-5.3

Zhipu AI

68.8

$1.4 in / $4.4 out