Skytells
HomeModelsCLIChangelog
  • Home
  • Models
  • CLI
  • Changelog
Skytells

Addressing the world's greatest challenges with AI. Enterprise research, foundation models, and infrastructure trusted by organizations worldwide since 2012.

Get Started

  • Console
  • Learn
  • Documentation
  • API Reference
  • Pricing
  • ModelsNew

Platform

  • Cloud AgentsNew
  • AI Solutions
  • Infrastructure
  • Edge Network
  • Trust Center
  • CLI

Resources

  • Blog
  • Changelog
  • AI Leaderboard
  • Research
  • Status

Company

  • About
  • Careers
  • Legal
  • Privacy Policy

© 2012–2026 Skytells, Inc. All rights reserved.

Live rankings

AI Model Leaderboard

Every major AI model ranked across benchmark quality, inference speed, agentic capability, programming aptitude, and cost efficiency — updated continuously from published evaluation data.

Explore full leaderboardBrowse model catalog

334

Tracked models

29

Providers

286

Benchmarked

29.3

Avg. index

OverallBenchmarksInferenceAgenticProgrammingValue / Price

334 models

RankModelProviderScoreBenchmarksInferenceAgenticProgrammingValuePrice
301

Gemma 2 27B

gemma-2-27b-it

textinference
Google

0.0

overall

0.00.00.00.00.0N/A
302

Gemma 2 9B

gemma-2-9b-it

textinference
Google

0.0

overall

0.00.00.00.00.0N/A
303

Gemma 3n E2B

gemma-3n-e2b

multimodalvisionmulti-input reasoning
Google

0.0

overall

0.00.00.00.00.0N/A
304

Gemma 3n E4B

gemma-3n-e4b

multimodalvisionmulti-input reasoning
Google

0.0

overall

0.00.00.00.00.0N/A
305

GLM-4.5V

glm-4.5v

multimodalvisionmulti-input reasoning
ZZhipu AI

0.0

overall

0.00.00.00.00.0N/A
306

Granite 3.3 8B Base

granite-3.3-8b-base

multimodalvisionmulti-input reasoning
IIBM

0.0

overall

0.00.00.00.00.0N/A
307

Granite 3.3 8B Instruct

granite-3.3-8b-instruct

multimodalvisionmulti-input reasoning
IIBM

0.0

overall

0.00.00.00.00.0N/A
308

IBM Granite 4.0 Tiny Preview

granite-4.0-tiny-preview

textinference
IIBM

0.0

overall

0.00.00.00.00.0N/A
309

Grok-2 Image 1212

grok-2-image-1212

textinference
xAI

0.0

overall

0.00.00.00.00.0N/A
310

Grok-4.1

grok-4.1-2025-11-17

multimodalvisionmulti-input reasoning
xAI

0.0

overall

0.00.00.00.00.0N/A
311

Grok-4.1 Thinking

grok-4.1-thinking-2025-11-17

multimodalvisionmulti-input reasoning
xAI

0.0

overall

0.00.00.00.00.0
312

Grok-4.20 Beta Non-Reasoning

grok-4.20-beta-0309-non-reasoning

multimodalvisionmulti-input reasoning
xAI

0.0

overall

0.00.00.00.00.0
313

Grok-4.20 Beta Reasoning

grok-4.20-beta-0309-reasoning

multimodalvisionmulti-input reasoning
xAI

0.0

overall

0.00.00.00.00.0
314

Grok-4.20 Multi-Agent Beta

grok-4.20-multi-agent-beta-0309

multimodalvisionmulti-input reasoning
xAI

0.0

overall

0.00.00.00.00.0
315

Grok-4 Fast Non-Reasoning

grok-4-fast-non-reasoning

multimodalvisionmulti-input reasoning
xAI

0.0

overall

0.00.00.00.00.0
316

Llama 3.1 Nemotron 70B Instruct

llama-3.1-nemotron-70b-instruct

textinference
NNVIDIA

0.0

overall

0.00.00.00.00.0N/A
317

MedGemma 4B IT

medgemma-4b-it

multimodalvisionmulti-input reasoning
Google

0.0

overall

0.00.00.00.00.0N/A
318

Ministral 3 (14B Base 2512)

ministral-3-14b-base-2512

multimodalvisionmulti-input reasoning
Mistral AI

0.0

overall

0.00.00.00.00.0
319

MiniStral 3 (14B Instruct 2512)

ministral-3-14b-instruct-2512

multimodalvisionmulti-input reasoning
Mistral AI

0.0

overall

0.00.00.00.00.0
320

Ministral 3 (3B Base 2512)

ministral-3-3b-base-2512

multimodalvisionmulti-input reasoning
Mistral AI

0.0

overall

0.00.00.00.00.0
301

Gemma 2 27B

Google

0.0

N/A

302

Gemma 2 9B

Google

0.0

N/A

303

Gemma 3n E2B

Google

0.0

N/A

304

Page 16 of 17 · 334 models

PreviousNext

Want benchmark charts, model comparison, and pricing analytics?

Sign in to access the full interactive leaderboard with deep benchmark breakdowns and model comparison tools.

Open full leaderboard

Rankings are based on multi-dimensional evaluation across benchmark quality, inference efficiency, and cost-per-output. Scores are updated continuously and may differ from individual third-party benchmarks.

N/A
N/A
N/A
N/A
N/A
N/A
N/A
N/A

Gemma 3n E4B

Google

0.0

N/A

305
Z

GLM-4.5V

Zhipu AI

0.0

N/A

306
I

Granite 3.3 8B Base

IBM

0.0

N/A

307
I

Granite 3.3 8B Instruct

IBM

0.0

N/A

308
I

IBM Granite 4.0 Tiny Preview

IBM

0.0

N/A

309

Grok-2 Image 1212

xAI

0.0

N/A

310

Grok-4.1

xAI

0.0

N/A

311

Grok-4.1 Thinking

xAI

0.0

N/A

312

Grok-4.20 Beta Non-Reasoning

xAI

0.0

N/A

313

Grok-4.20 Beta Reasoning

xAI

0.0

N/A

314

Grok-4.20 Multi-Agent Beta

xAI

0.0

N/A

315

Grok-4 Fast Non-Reasoning

xAI

0.0

N/A

316
N

Llama 3.1 Nemotron 70B Instruct

NVIDIA

0.0

N/A

317

MedGemma 4B IT

Google

0.0

N/A

318

Ministral 3 (14B Base 2512)

Mistral AI

0.0

N/A

319

MiniStral 3 (14B Instruct 2512)

Mistral AI

0.0

N/A

320

Ministral 3 (3B Base 2512)

Mistral AI

0.0

N/A