Skytells
HomeModelsCLIChangelog
  • Home
  • Models
  • CLI
  • Changelog
Skytells

Addressing the world's greatest challenges with AI. Enterprise research, foundation models, and infrastructure trusted by organizations worldwide since 2012.

Get Started

  • Console
  • Learn
  • Documentation
  • API Reference
  • Pricing
  • ModelsNew

Platform

  • Cloud AgentsNew
  • AI Solutions
  • Infrastructure
  • Edge Network
  • Trust Center
  • CLI

Resources

  • Blog
  • Changelog
  • AI Leaderboard
  • Research
  • Status

Company

  • About
  • Careers
  • Legal
  • Privacy Policy

© 2012–2026 Skytells, Inc. All rights reserved.

Live rankings

AI Model Leaderboard

Every major AI model ranked across benchmark quality, inference speed, agentic capability, programming aptitude, and cost efficiency — updated continuously from published evaluation data.

Explore full leaderboardBrowse model catalog

334

Tracked models

29

Providers

286

Benchmarked

29.3

Avg. index

OverallBenchmarksInferenceAgenticProgrammingValue / Price

334 models

RankModelProviderScoreBenchmarksInferenceAgenticProgrammingValuePrice
281

Jamba 1.5 Mini

jamba-1.5-mini

textinference
AAI21 Labs

4.3

overall

4.30.00.00.00.0N/A
282

GPT-5.1 Codex Mini

gpt-5.1-codex-mini

multimodalvisionmulti-input reasoning
OpenAI

3.8

overall

3.80.00.00.00.0
283

Llama 3.2 11B Instruct

llama-3.2-11b-instruct

multimodalvisionmulti-input reasoning
MMeta

3.8

overall

3.80.00.00.00.0N/A
284

Gemini 1.0 Pro

gemini-1.0-pro

multimodalvisionmulti-input reasoning
Google

3.0

overall

3.00.00.00.00.0N/A
285

Llama 3.1 8B Instruct

llama-3.1-8b-instruct

textinference
MMeta

3.0

overall

3.00.00.00.00.0N/A
286

Phi-3.5-mini-instruct

phi-3.5-mini-instruct

multimodalvisionmulti-input reasoning
MMicrosoft

2.4

overall

2.40.00.00.00.0N/A
287

Phi-3.5-vision-instruct

phi-3.5-vision-instruct

multimodalvisionmulti-input reasoning
MMicrosoft

2.3

overall

2.30.00.00.00.0N/A
288

Qwen2 7B Instruct

qwen2-7b-instruct

textinference
AAlibaba Cloud / Qwen Team

2.2

overall

2.20.00.00.00.0N/A
289

Phi 4 Mini

phi-4-mini

textinference
MMicrosoft

1.9

overall

1.90.00.00.00.0N/A
290

Gemma 3n E4B Instructed

gemma-3n-e4b-it

multimodalvisionmulti-input reasoning
Google

1.2

overall

1.20.00.00.00.0
291

Gemma 3n E4B Instructed LiteRT Preview

gemma-3n-e4b-it-litert-preview

multimodalvisionmulti-input reasoning
Google

1.2

overall

1.20.00.00.00.0
292

DeepSeek VL2 Tiny

deepseek-vl2-tiny

multimodalvisionmulti-input reasoning
DeepSeek

1.1

overall

1.10.00.00.00.0
293

Gemma 3n E2B Instructed

gemma-3n-e2b-it

multimodalvisionmulti-input reasoning
Google

1.0

overall

1.00.00.00.00.0
294

Gemma 3n E2B Instructed LiteRT (Preview)

gemma-3n-e2b-it-litert-preview

multimodalvisionmulti-input reasoning
Google

1.0

overall

1.00.00.00.00.0
295

Gemma 3 1B

gemma-3-1b-it

textinference
Google

0.9

overall

0.90.00.00.00.0N/A
296

DeepSeek-V2.5

deepseek-v2.5

codeprogrammingtool use
DeepSeek

0.7

overall

0.00.00.00.70.0N/A
297

Codestral-22B

codestral-22b

textinference
Mistral AI

0.0

overall

0.00.00.00.00.0N/A
298

Command R+

command-r-plus-04-2024

textinference
Cohere

0.0

overall

0.00.00.00.00.0N/A
299

DeepSeek-R1

deepseek-r1

textinference
DeepSeek

0.0

overall

0.00.00.00.00.0N/A
300

Gemini 3.5 Flash Cyber

gemini-3.5-flash-cyber

textinference
Google

0.0

overall

0.00.00.00.00.0N/A
281
A

Jamba 1.5 Mini

AI21 Labs

4.3

N/A

282

GPT-5.1 Codex Mini

OpenAI

3.8

N/A

283
M

Llama 3.2 11B Instruct

Meta

3.8

N/A

284

Page 15 of 17 · 334 models

PreviousNext

Want benchmark charts, model comparison, and pricing analytics?

Sign in to access the full interactive leaderboard with deep benchmark breakdowns and model comparison tools.

Open full leaderboard

Rankings are based on multi-dimensional evaluation across benchmark quality, inference efficiency, and cost-per-output. Scores are updated continuously and may differ from individual third-party benchmarks.

N/A
N/A
N/A
N/A
N/A
N/A

Gemini 1.0 Pro

Google

3.0

N/A

285
M

Llama 3.1 8B Instruct

Meta

3.0

N/A

286
M

Phi-3.5-mini-instruct

Microsoft

2.4

N/A

287
M

Phi-3.5-vision-instruct

Microsoft

2.3

N/A

288
A

Qwen2 7B Instruct

Alibaba Cloud / Qwen Team

2.2

N/A

289
M

Phi 4 Mini

Microsoft

1.9

N/A

290

Gemma 3n E4B Instructed

Google

1.2

N/A

291

Gemma 3n E4B Instructed LiteRT Preview

Google

1.2

N/A

292

DeepSeek VL2 Tiny

DeepSeek

1.1

N/A

293

Gemma 3n E2B Instructed

Google

1.0

N/A

294

Gemma 3n E2B Instructed LiteRT (Preview)

Google

1.0

N/A

295

Gemma 3 1B

Google

0.9

N/A

296

DeepSeek-V2.5

DeepSeek

0.7

N/A

297

Codestral-22B

Mistral AI

0.0

N/A

298

Command R+

Cohere

0.0

N/A

299

DeepSeek-R1

DeepSeek

0.0

N/A

300

Gemini 3.5 Flash Cyber

Google

0.0

N/A