Skytells
HomeModelsCLIChangelog
  • Home
  • Models
  • CLI
  • Changelog
Skytells

Addressing the world's greatest challenges with AI. Enterprise research, foundation models, and infrastructure trusted by organizations worldwide since 2012.

Get Started

  • Console
  • Learn
  • Documentation
  • API Reference
  • Pricing
  • ModelsNew

Platform

  • Cloud AgentsNew
  • AI Solutions
  • Infrastructure
  • Edge Network
  • Trust Center
  • CLI

Resources

  • Blog
  • Changelog
  • AI Leaderboard
  • Research
  • Status

Company

  • About
  • Careers
  • Legal
  • Privacy Policy

© 2012–2026 Skytells, Inc. All rights reserved.

Live rankings

AI Model Leaderboard

Every major AI model ranked across benchmark quality, inference speed, agentic capability, programming aptitude, and cost efficiency — updated continuously from published evaluation data.

Explore full leaderboardBrowse model catalog

334

Tracked models

29

Providers

286

Benchmarked

15.1

Avg. index

OverallBenchmarksInferenceAgenticProgrammingValue / Price

334 models

RankModelProviderScoreBenchmarksInferenceAgenticProgrammingValuePrice
221

Grok-4 Fast Non-Reasoning

grok-4-fast-non-reasoning

multimodalvisionmulti-input reasoning
xAI

0.0

Programming

0.00.00.00.00.0N/A
222

Grok-4 Fast Reasoning

grok-4-fast-reasoning

multimodalvisionmulti-input reasoning
xAI

0.0

Programming

0.063.00.00.077.3
223

Grok-4 Heavy

grok-4-heavy

multimodalvisionmulti-input reasoning
xAI

0.0

Programming

69.50.00.00.00.0N/A
224

Hermes 3 70B

hermes-3-70b

textinference
NNous Research

0.0

Programming

27.60.00.00.00.0N/A
225

Jamba 1.5 Large

jamba-1.5-large

textinference
AAI21 Labs

0.0

Programming

7.50.00.00.00.0N/A
226

Jamba 1.5 Mini

jamba-1.5-mini

textinference
AAI21 Labs

0.0

Programming

4.30.00.00.00.0N/A
227

K-EXAONE-236B-A23B

k-exaone-236b-a23b

multimodalvisionmulti-input reasoning
LLG AI Research

0.0

Programming

43.90.00.00.00.0N/A
228

Kimi-k1.5

kimi-k1.5

multimodalvisionmulti-input reasoning
Moonshot AI

0.0

Programming

34.70.00.00.00.0N/A
229

Kimi K2 0905

kimi-k2-0905

textinference
Moonshot AI

0.0

Programming

41.00.00.00.00.0N/A
230

Kimi K2.7 Code

kimi-k2.7-code

multimodalvisionmulti-input reasoning
Moonshot AI

0.0

Programming

0.032.947.20.047.6
231

Kimi K2 Base

kimi-k2-base

textinference
Moonshot AI

0.0

Programming

26.00.00.00.00.0N/A
232

Kimi K3

kimi-k3

multimodalvisionmulti-input reasoning
Moonshot AI

0.0

Programming

76.184.884.20.012.2$3 in / $15 out
233

Llama 3.1 405B Instruct

llama-3.1-405b-instruct

textinference
MMeta

0.0

Programming

18.30.00.00.00.0N/A
234

Llama 3.1 70B Instruct

llama-3.1-70b-instruct

textinference
MMeta

0.0

Programming

10.30.00.00.00.0N/A
235

Llama 3.1 8B Instruct

llama-3.1-8b-instruct

textinference
MMeta

0.0

Programming

3.00.00.00.00.0N/A
236

Llama 3.1 Nemotron 70B Instruct

llama-3.1-nemotron-70b-instruct

textinference
NNVIDIA

0.0

Programming

0.00.00.00.00.0N/A
237

Llama 3.1 Nemotron Nano 8B V1

llama-3.1-nemotron-nano-8b-v1

textinference
NNVIDIA

0.0

Programming

15.00.00.00.00.0N/A
238

Llama 3.1 Nemotron Ultra 253B v1

llama-3.1-nemotron-ultra-253b-v1

textinference
NNVIDIA

0.0

Programming

33.00.00.00.00.0N/A
239

Llama 3.2 11B Instruct

llama-3.2-11b-instruct

multimodalvisionmulti-input reasoning
MMeta

0.0

Programming

3.80.00.00.00.0N/A
240

Llama 3.2 3B Instruct

llama-3.2-3b-instruct

textinference
MMeta

0.0

Programming

4.80.00.00.00.0N/A
221

Grok-4 Fast Non-Reasoning

xAI

0.0

N/A

222

Grok-4 Fast Reasoning

xAI

0.0

$0.2 in / $0.5 out

223

Grok-4 Heavy

xAI

0.0

N/A

224

Page 12 of 17 · 334 models

PreviousNext

Want benchmark charts, model comparison, and pricing analytics?

Sign in to access the full interactive leaderboard with deep benchmark breakdowns and model comparison tools.

Open full leaderboard

Rankings are based on multi-dimensional evaluation across benchmark quality, inference efficiency, and cost-per-output. Scores are updated continuously and may differ from individual third-party benchmarks.

$0.2 in / $0.5 out
$0.74 in / $3.5 out
N

Hermes 3 70B

Nous Research

0.0

N/A

225
A

Jamba 1.5 Large

AI21 Labs

0.0

N/A

226
A

Jamba 1.5 Mini

AI21 Labs

0.0

N/A

227
L

K-EXAONE-236B-A23B

LG AI Research

0.0

N/A

228

Kimi-k1.5

Moonshot AI

0.0

N/A

229

Kimi K2 0905

Moonshot AI

0.0

N/A

230

Kimi K2.7 Code

Moonshot AI

0.0

$0.74 in / $3.5 out

231

Kimi K2 Base

Moonshot AI

0.0

N/A

232

Kimi K3

Moonshot AI

0.0

$3 in / $15 out

233
M

Llama 3.1 405B Instruct

Meta

0.0

N/A

234
M

Llama 3.1 70B Instruct

Meta

0.0

N/A

235
M

Llama 3.1 8B Instruct

Meta

0.0

N/A

236
N

Llama 3.1 Nemotron 70B Instruct

NVIDIA

0.0

N/A

237
N

Llama 3.1 Nemotron Nano 8B V1

NVIDIA

0.0

N/A

238
N

Llama 3.1 Nemotron Ultra 253B v1

NVIDIA

0.0

N/A

239
M

Llama 3.2 11B Instruct

Meta

0.0

N/A

240
M

Llama 3.2 3B Instruct

Meta

0.0

N/A