Skytells
HomeModelsCLIChangelog
  • Home
  • Models
  • CLI
  • Changelog
Skytells

Addressing the world's greatest challenges with AI. Enterprise research, foundation models, and infrastructure trusted by organizations worldwide since 2012.

Get Started

  • Console
  • Learn
  • Documentation
  • API Reference
  • Pricing
  • ModelsNew

Platform

  • Cloud AgentsNew
  • AI Solutions
  • Infrastructure
  • Edge Network
  • Trust Center
  • CLI

Resources

  • Blog
  • Changelog
  • AI Leaderboard
  • Research
  • Status

Company

  • About
  • Careers
  • Legal
  • Privacy Policy

© 2012–2026 Skytells, Inc. All rights reserved.

Live rankings

AI Model Leaderboard

Every major AI model ranked across benchmark quality, inference speed, agentic capability, programming aptitude, and cost efficiency — updated continuously from published evaluation data.

Explore full leaderboardBrowse model catalog

334

Tracked models

29

Providers

286

Benchmarked

12.6

Avg. index

OverallBenchmarksInferenceAgenticProgrammingValue / Price

334 models

RankModelProviderScoreBenchmarksInferenceAgenticProgrammingValuePrice
1

Nemotron 3 Nano (30B A3B)

nemotron-3-nano-30b-a3b

codeprogrammingtool use
NNVIDIA

100.0

Value / Price

43.632.93.03.8100.0$0.06 in / $0.24 out
2

DeepSeek-V4-Flash-Max

deepseek-v4-flash-max

codeprogrammingtool use
DeepSeek

98.8

Value / Price

56.284.835.341.498.8
3

LongCat-Flash-Lite

longcat-flash-lite

codeprogrammingtool use
Meituan

95.6

Value / Price

22.872.830.123.995.6
4

GPT-4.1 nano

gpt-4.1-nano-2025-04-14

multimodalvisionmulti-input reasoning
OpenAI

94.9

Value / Price

11.687.80.00.094.9
5

Step-3.5-Flash

step-3.5-flash

codeprogrammingtool use
SStepFun

93.9

Value / Price

62.859.836.548.693.9$0.1 in / $0.4 out
6

MiMo-V2.5

mimo-v2.5

multimodalvisionmulti-input reasoning
Xiaomi

92.7

Value / Price

47.784.80.027.292.7$0.168 in / $0.336 out
7

Gemma 4 31B

gemma-4-31b-it

multimodalvisionmulti-input reasoning
Google

91.5

Value / Price

55.432.90.00.091.5
8

Gemma 4 26B-A4B

gemma-4-26b-a4b-it

multimodalvisionmulti-input reasoning
Google

90.2

Value / Price

43.832.90.00.090.2
9

Qwen3 VL 4B Instruct

qwen3-vl-4b-instruct

multimodalvisionmulti-input reasoning
AAlibaba Cloud / Qwen Team

85.4

Value / Price

18.232.917.70.085.4
10

Mercury 2

mercury-2

codeprogrammingtool use
IInception

84.4

Value / Price

42.768.90.015.384.4$0.25 in / $0.75 out
11

DeepSeek-V3.2 (Non-thinking)

deepseek-chat

textinference
DeepSeek

82.7

Value / Price

0.052.00.00.082.7$0.28 in / $0.42 out
12

Mistral Small 4

mistral-small-latest

multimodalvisionmulti-input reasoning
Mistral AI

81.7

Value / Price

31.323.20.00.081.7
13

Qwen3 VL 4B Thinking

qwen3-vl-4b-thinking

multimodalvisionmulti-input reasoning
AAlibaba Cloud / Qwen Team

79.3

Value / Price

20.232.917.00.079.3
14

Qwen3 30B A3B

qwen3-30b-a3b

textinference
AAlibaba Cloud / Qwen Team

78.5

Value / Price

23.726.60.00.078.5$0.1 in / $0.44 out
15

MiMo-V2.5-Pro

mimo-v2.5-pro

codeprogrammingtool use
Xiaomi

78.0

Value / Price

36.284.80.056.378.0$0.435 in / $0.87 out
16

Qwen3 32B

qwen3-32b

textinference
AAlibaba Cloud / Qwen Team

78.0

Value / Price

20.12.20.00.078.0$0.1 in / $0.3 out
17

Grok-4.1 Fast Non-Reasoning

grok-4-1-fast-non-reasoning

multimodalvisionmulti-input reasoning
xAI

77.3

Value / Price

0.063.00.00.077.3
18

Grok-4.1 Fast Reasoning

grok-4-1-fast-reasoning

multimodalvisionmulti-input reasoning
xAI

77.3

Value / Price

0.063.00.00.077.3
19

Grok-4 Fast Reasoning

grok-4-fast-reasoning

multimodalvisionmulti-input reasoning
xAI

77.3

Value / Price

0.063.00.00.077.3
20

GPT-5.4 nano

gpt-5.4-nano

multimodalvisionmulti-input reasoning
OpenAI

76.8

Value / Price

41.844.56.98.276.8
1
N

Nemotron 3 Nano (30B A3B)

NVIDIA

100.0

$0.06 in / $0.24 out

2

DeepSeek-V4-Flash-Max

DeepSeek

98.8

$0.1 in / $0.2 out

3

LongCat-Flash-Lite

Meituan

95.6

$0.1 in / $0.4 out

Page 1 of 17 · 334 models

Next

Want benchmark charts, model comparison, and pricing analytics?

Sign in to access the full interactive leaderboard with deep benchmark breakdowns and model comparison tools.

Open full leaderboard

Rankings are based on multi-dimensional evaluation across benchmark quality, inference efficiency, and cost-per-output. Scores are updated continuously and may differ from individual third-party benchmarks.

$0.1 in / $0.2 out
$0.1 in / $0.4 out
$0.1 in / $0.4 out
$0.13 in / $0.38 out
$0.13 in / $0.4 out
$0.1 in / $0.6 out
$0.15 in / $0.6 out
$0.1 in / $1 out
$0.2 in / $0.5 out
$0.2 in / $0.5 out
$0.2 in / $0.5 out
$0.2 in / $1.25 out
4

GPT-4.1 nano

OpenAI

94.9

$0.1 in / $0.4 out

5
S

Step-3.5-Flash

StepFun

93.9

$0.1 in / $0.4 out

6

MiMo-V2.5

Xiaomi

92.7

$0.168 in / $0.336 out

7

Gemma 4 31B

Google

91.5

$0.13 in / $0.38 out

8

Gemma 4 26B-A4B

Google

90.2

$0.13 in / $0.4 out

9
A

Qwen3 VL 4B Instruct

Alibaba Cloud / Qwen Team

85.4

$0.1 in / $0.6 out

10
I

Mercury 2

Inception

84.4

$0.25 in / $0.75 out

11

DeepSeek-V3.2 (Non-thinking)

DeepSeek

82.7

$0.28 in / $0.42 out

12

Mistral Small 4

Mistral AI

81.7

$0.15 in / $0.6 out

13
A

Qwen3 VL 4B Thinking

Alibaba Cloud / Qwen Team

79.3

$0.1 in / $1 out

14
A

Qwen3 30B A3B

Alibaba Cloud / Qwen Team

78.5

$0.1 in / $0.44 out

15

MiMo-V2.5-Pro

Xiaomi

78.0

$0.435 in / $0.87 out

16
A

Qwen3 32B

Alibaba Cloud / Qwen Team

78.0

$0.1 in / $0.3 out

17

Grok-4.1 Fast Non-Reasoning

xAI

77.3

$0.2 in / $0.5 out

18

Grok-4.1 Fast Reasoning

xAI

77.3

$0.2 in / $0.5 out

19

Grok-4 Fast Reasoning

xAI

77.3

$0.2 in / $0.5 out

20

GPT-5.4 nano

OpenAI

76.8

$0.2 in / $1.25 out