Skytells
HomeModelsCLIChangelog
  • Home
  • Models
  • CLI
  • Changelog
Skytells

Addressing the world's greatest challenges with AI. Enterprise research, foundation models, and infrastructure trusted by organizations worldwide since 2012.

Get Started

  • Console
  • Learn
  • Documentation
  • API Reference
  • Pricing
  • ModelsNew

Platform

  • Cloud AgentsNew
  • AI Solutions
  • Infrastructure
  • Edge Network
  • Trust Center
  • CLI

Resources

  • Blog
  • Changelog
  • AI Leaderboard
  • Research
  • Status

Company

  • About
  • Careers
  • Legal
  • Privacy Policy

© 2012–2026 Skytells, Inc. All rights reserved.

Live rankings

AI Model Leaderboard

Every major AI model ranked across benchmark quality, inference speed, agentic capability, programming aptitude, and cost efficiency — updated continuously from published evaluation data.

Explore full leaderboardBrowse model catalog

334

Tracked models

29

Providers

286

Benchmarked

29.3

Avg. index

OverallBenchmarksInferenceAgenticProgrammingValue / Price

334 models

RankModelProviderScoreBenchmarksInferenceAgenticProgrammingValuePrice
61

GPT OSS 20B High

gpt-oss-20b-high

textinference
OpenAI

52.3

overall

52.30.00.00.00.0N/A
62

Seed 2.0 Lite

seed-2.0-lite

multimodalvisionmulti-input reasoning
BByteDance

51.9

overall

56.50.00.046.10.0N/A
63

Seed 2.0 Pro

seed-2.0-pro

multimodalvisionmulti-input reasoning
BByteDance

51.0

overall

66.823.244.856.054.9$0.5 in / $3 out
64

Grok-3 Mini

grok-3-mini

multimodalvisionmulti-input reasoning
xAI

51.0

overall

51.00.00.00.00.0N/A
65

MiMo-V2-Omni

mimo-v2-omni

multimodalvisionmulti-input reasoning
Xiaomi

50.9

overall

0.00.00.050.90.0N/A
66

Claude Sonnet 4.5

claude-sonnet-4-5-20250929

multimodalvisionmulti-input reasoning
Anthropic

50.8

overall

51.412.669.974.612.0
67

MiniMax M2.1

minimax-m2.1

codeprogrammingtool use
MiniMax

50.7

overall

39.168.945.747.672.9$0.3 in / $1.2 out
68

Nova 2 Pro

nova-2-pro

multimodalvisionmulti-input reasoning
AAmazon

50.4

overall

45.30.057.249.60.0N/A
69

GPT-5.4

gpt-5.4

texttext-to-textlanguage
OpenAI

50.3

overall

70.537.248.950.318.5
70

Kimi K2.5

kimi-k2.5

multimodalvisionmulti-input reasoning
Moonshot AI

49.8

overall

63.70.041.041.80.0N/A
71

GPT-5 Codex

gpt-5-codex-2025-09-15

codeprogrammingtool use
OpenAI

49.8

overall

0.00.00.049.80.0N/A
72

Grok-3

grok-3

multimodalvisionmulti-input reasoning
xAI

49.6

overall

58.450.20.00.024.1$3 in / $15 out
73

Gemma 4 26B-A4B

gemma-4-26b-a4b-it

multimodalvisionmulti-input reasoning
Google

49.2

overall

43.832.90.00.090.2
74

Grok-4

grok-4

multimodalvisionmulti-input reasoning
xAI

49.1

overall

49.10.00.00.00.0N/A
75

GLM-5V-Turbo

glm-5v-turbo

multimodalvisionmulti-input reasoning
ZZhipu AI

49.1

overall

0.00.049.10.00.0N/A
76

Qwen3.5-122B-A10B

qwen3.5-122b-a10b

multimodalvisionmulti-input reasoning
AAlibaba Cloud / Qwen Team

48.8

overall

60.50.044.638.30.0N/A
77

MAI-Thinking-1

mai-thinking-1

codeprogrammingtool use
MMicrosoft

47.8

overall

60.10.00.032.20.0N/A
78

GPT-5.5 Instant

gpt-5.5-instant

multimodalvisionmulti-input reasoning
OpenAI

47.5

overall

49.562.80.00.017.3
79

GPT-5.1 Codex

gpt-5.1-codex

multimodalvisionmulti-input reasoning
OpenAI

47.2

overall

0.00.00.047.20.0N/A
80

Step3-VL-10B

step3-vl-10b

multimodalvisionmulti-input reasoning
SStepFun

46.9

overall

46.90.00.00.00.0N/A
61

GPT OSS 20B High

OpenAI

52.3

N/A

62
B

Seed 2.0 Lite

ByteDance

51.9

N/A

63
B

Seed 2.0 Pro

ByteDance

51.0

$0.5 in / $3 out

64

Page 4 of 17 · 334 models

PreviousNext

Want benchmark charts, model comparison, and pricing analytics?

Sign in to access the full interactive leaderboard with deep benchmark breakdowns and model comparison tools.

Open full leaderboard

Rankings are based on multi-dimensional evaluation across benchmark quality, inference efficiency, and cost-per-output. Scores are updated continuously and may differ from individual third-party benchmarks.

$3 in / $15 out
$2.5 in / $15 out
$0.13 in / $0.4 out
$5 in / $30 out

Grok-3 Mini

xAI

51.0

N/A

65

MiMo-V2-Omni

Xiaomi

50.9

N/A

66

Claude Sonnet 4.5

Anthropic

50.8

$3 in / $15 out

67

MiniMax M2.1

MiniMax

50.7

$0.3 in / $1.2 out

68
A

Nova 2 Pro

Amazon

50.4

N/A

69

GPT-5.4

OpenAI

50.3

$2.5 in / $15 out

70

Kimi K2.5

Moonshot AI

49.8

N/A

71

GPT-5 Codex

OpenAI

49.8

N/A

72

Grok-3

xAI

49.6

$3 in / $15 out

73

Gemma 4 26B-A4B

Google

49.2

$0.13 in / $0.4 out

74

Grok-4

xAI

49.1

N/A

75
Z

GLM-5V-Turbo

Zhipu AI

49.1

N/A

76
A

Qwen3.5-122B-A10B

Alibaba Cloud / Qwen Team

48.8

N/A

77
M

MAI-Thinking-1

Microsoft

47.8

N/A

78

GPT-5.5 Instant

OpenAI

47.5

$5 in / $30 out

79

GPT-5.1 Codex

OpenAI

47.2

N/A

80
S

Step3-VL-10B

StepFun

46.9

N/A