Skytells
HomeModelsCLIChangelog
  • Home
  • Models
  • CLI
  • Changelog
Skytells

Addressing the world's greatest challenges with AI. Enterprise research, foundation models, and infrastructure trusted by organizations worldwide since 2012.

Get Started

  • Console
  • Learn
  • Documentation
  • API Reference
  • Pricing
  • ModelsNew

Platform

  • Cloud AgentsNew
  • AI Solutions
  • Infrastructure
  • Edge Network
  • Trust Center
  • CLI

Resources

  • Blog
  • Changelog
  • AI Leaderboard
  • Research
  • Status

Company

  • About
  • Careers
  • Legal
  • Privacy Policy

© 2012–2026 Skytells, Inc. All rights reserved.

Live rankings

AI Model Leaderboard

Every major AI model ranked across benchmark quality, inference speed, agentic capability, programming aptitude, and cost efficiency — updated continuously from published evaluation data.

Explore full leaderboardBrowse model catalog

334

Tracked models

29

Providers

286

Benchmarked

12.2

Avg. index

OverallBenchmarksInferenceAgenticProgrammingValue / Price

334 models

RankModelProviderScoreBenchmarksInferenceAgenticProgrammingValuePrice
221

Grok-4.20 Multi-Agent Beta

grok-4.20-multi-agent-beta-0309

multimodalvisionmulti-input reasoning
xAI

0.0

Agentic

0.00.00.00.00.0N/A
222

Grok 4.3

grok-4.3

textinference
xAI

0.0

Agentic

0.062.80.00.052.4$1.25 in / $2.5 out
223

Grok 4.5

grok-4.5

multimodalvisionmulti-input reasoning
xAI

0.0

Agentic

69.338.20.070.835.6$2 in / $6 out
224

Grok-4 Fast Non-Reasoning

grok-4-fast-non-reasoning

multimodalvisionmulti-input reasoning
xAI

0.0

Agentic

0.00.00.00.00.0
225

Grok-4 Fast Reasoning

grok-4-fast-reasoning

multimodalvisionmulti-input reasoning
xAI

0.0

Agentic

0.063.00.00.077.3
226

Grok-4 Heavy

grok-4-heavy

multimodalvisionmulti-input reasoning
xAI

0.0

Agentic

69.50.00.00.00.0N/A
227

Grok Code Fast 1

grok-code-fast-1

codeprogrammingtool use
xAI

0.0

Agentic

0.027.20.036.160.5$0.2 in / $1.5 out
228

Hermes 3 70B

hermes-3-70b

textinference
NNous Research

0.0

Agentic

27.60.00.00.00.0N/A
229

Jamba 1.5 Large

jamba-1.5-large

textinference
AAI21 Labs

0.0

Agentic

7.50.00.00.00.0N/A
230

Jamba 1.5 Mini

jamba-1.5-mini

textinference
AAI21 Labs

0.0

Agentic

4.30.00.00.00.0N/A
231

K-EXAONE-236B-A23B

k-exaone-236b-a23b

multimodalvisionmulti-input reasoning
LLG AI Research

0.0

Agentic

43.90.00.00.00.0N/A
232

Kimi-k1.5

kimi-k1.5

multimodalvisionmulti-input reasoning
Moonshot AI

0.0

Agentic

34.70.00.00.00.0N/A
233

Kimi K2 0905

kimi-k2-0905

textinference
Moonshot AI

0.0

Agentic

41.00.00.00.00.0N/A
234

Kimi K2 Base

kimi-k2-base

textinference
Moonshot AI

0.0

Agentic

26.00.00.00.00.0N/A
235

Llama 3.1 405B Instruct

llama-3.1-405b-instruct

textinference
MMeta

0.0

Agentic

18.30.00.00.00.0N/A
236

Llama 3.1 70B Instruct

llama-3.1-70b-instruct

textinference
MMeta

0.0

Agentic

10.30.00.00.00.0N/A
237

Llama 3.1 8B Instruct

llama-3.1-8b-instruct

textinference
MMeta

0.0

Agentic

3.00.00.00.00.0N/A
238

Llama 3.1 Nemotron 70B Instruct

llama-3.1-nemotron-70b-instruct

textinference
NNVIDIA

0.0

Agentic

0.00.00.00.00.0N/A
239

Llama 3.1 Nemotron Nano 8B V1

llama-3.1-nemotron-nano-8b-v1

textinference
NNVIDIA

0.0

Agentic

15.00.00.00.00.0N/A
240

Llama 3.1 Nemotron Ultra 253B v1

llama-3.1-nemotron-ultra-253b-v1

textinference
NNVIDIA

0.0

Agentic

33.00.00.00.00.0N/A
221

Grok-4.20 Multi-Agent Beta

xAI

0.0

N/A

222

Grok 4.3

xAI

0.0

$1.25 in / $2.5 out

223

Grok 4.5

xAI

0.0

$2 in / $6 out

224

Page 12 of 17 · 334 models

PreviousNext

Want benchmark charts, model comparison, and pricing analytics?

Sign in to access the full interactive leaderboard with deep benchmark breakdowns and model comparison tools.

Open full leaderboard

Rankings are based on multi-dimensional evaluation across benchmark quality, inference efficiency, and cost-per-output. Scores are updated continuously and may differ from individual third-party benchmarks.

N/A
$0.2 in / $0.5 out

Grok-4 Fast Non-Reasoning

xAI

0.0

N/A

225

Grok-4 Fast Reasoning

xAI

0.0

$0.2 in / $0.5 out

226

Grok-4 Heavy

xAI

0.0

N/A

227

Grok Code Fast 1

xAI

0.0

$0.2 in / $1.5 out

228
N

Hermes 3 70B

Nous Research

0.0

N/A

229
A

Jamba 1.5 Large

AI21 Labs

0.0

N/A

230
A

Jamba 1.5 Mini

AI21 Labs

0.0

N/A

231
L

K-EXAONE-236B-A23B

LG AI Research

0.0

N/A

232

Kimi-k1.5

Moonshot AI

0.0

N/A

233

Kimi K2 0905

Moonshot AI

0.0

N/A

234

Kimi K2 Base

Moonshot AI

0.0

N/A

235
M

Llama 3.1 405B Instruct

Meta

0.0

N/A

236
M

Llama 3.1 70B Instruct

Meta

0.0

N/A

237
M

Llama 3.1 8B Instruct

Meta

0.0

N/A

238
N

Llama 3.1 Nemotron 70B Instruct

NVIDIA

0.0

N/A

239
N

Llama 3.1 Nemotron Nano 8B V1

NVIDIA

0.0

N/A

240
N

Llama 3.1 Nemotron Ultra 253B v1

NVIDIA

0.0

N/A