Skytells
HomeModelsCLIChangelog
  • Home
  • Models
  • CLI
  • Changelog
Skytells

Addressing the world's greatest challenges with AI. Enterprise research, foundation models, and infrastructure trusted by organizations worldwide since 2012.

Get Started

  • Console
  • Learn
  • Documentation
  • API Reference
  • Pricing
  • ModelsNew

Platform

  • Cloud AgentsNew
  • AI Solutions
  • Infrastructure
  • Edge Network
  • Trust Center
  • CLI

Resources

  • Blog
  • Changelog
  • AI Leaderboard
  • Research
  • Status

Company

  • About
  • Careers
  • Legal
  • Privacy Policy

© 2012–2026 Skytells, Inc. All rights reserved.

Live rankings

AI Model Leaderboard

Every major AI model ranked across benchmark quality, inference speed, agentic capability, programming aptitude, and cost efficiency — updated continuously from published evaluation data.

Explore full leaderboardBrowse model catalog

334

Tracked models

29

Providers

286

Benchmarked

12.2

Avg. index

OverallBenchmarksInferenceAgenticProgrammingValue / Price

334 models

RankModelProviderScoreBenchmarksInferenceAgenticProgrammingValuePrice
241

Llama 3.2 11B Instruct

llama-3.2-11b-instruct

multimodalvisionmulti-input reasoning
MMeta

0.0

Agentic

3.80.00.00.00.0N/A
242

Llama 3.2 3B Instruct

llama-3.2-3b-instruct

textinference
MMeta

0.0

Agentic

4.80.00.00.00.0N/A
243

Llama 3.2 90B Instruct

llama-3.2-90b-instruct

multimodalvisionmulti-input reasoning
MMeta

0.0

Agentic

14.90.00.00.00.0N/A
244

Llama 3.3 70B Instruct

llama-3.3-70b-instruct

textinference
MMeta

0.0

Agentic

18.00.00.00.00.0N/A
245

Llama-3.3 Nemotron Super 49B v1

llama-3.3-nemotron-super-49b-v1

textinference
NNVIDIA

0.0

Agentic

21.30.00.00.00.0N/A
246

Llama 4 Maverick

llama-4-maverick

multimodalvisionmulti-input reasoning
MMeta

0.0

Agentic

32.60.00.00.00.0N/A
247

Llama 4 Scout

llama-4-scout

multimodalvisionmulti-input reasoning
MMeta

0.0

Agentic

27.60.00.00.00.0N/A
248

LongCat-Flash-Thinking

longcat-flash-thinking

codeprogrammingtool use
Meituan

0.0

Agentic

48.20.00.018.40.0
249

Magistral Medium

magistral-medium

multimodalvisionmulti-input reasoning
Mistral AI

0.0

Agentic

20.60.00.00.00.0
250

Magistral Small 2506

magistral-small-2506

textinference
Mistral AI

0.0

Agentic

22.70.00.00.00.0N/A
251

MAI-Code-1-Flash

mai-code-1-flash

codeprogrammingtool use
MMicrosoft

0.0

Agentic

32.20.00.021.40.0N/A
252

MAI-Thinking-1

mai-thinking-1

codeprogrammingtool use
MMicrosoft

0.0

Agentic

60.10.00.032.20.0N/A
253

MedGemma 4B IT

medgemma-4b-it

multimodalvisionmulti-input reasoning
Google

0.0

Agentic

0.00.00.00.00.0N/A
254

Mercury 2

mercury-2

codeprogrammingtool use
IInception

0.0

Agentic

42.768.90.015.384.4$0.25 in / $0.75 out
255

MiMo-V2.5

mimo-v2.5

multimodalvisionmulti-input reasoning
Xiaomi

0.0

Agentic

47.784.80.027.292.7$0.168 in / $0.336 out
256

MiMo-V2.5-Pro

mimo-v2.5-pro

codeprogrammingtool use
Xiaomi

0.0

Agentic

36.284.80.056.378.0$0.435 in / $0.87 out
257

MiMo-V2-Omni

mimo-v2-omni

multimodalvisionmulti-input reasoning
Xiaomi

0.0

Agentic

0.00.00.050.90.0N/A
258

MiMo-V2-Pro

mimo-v2-pro

codeprogrammingtool use
Xiaomi

0.0

Agentic

0.00.00.061.90.0N/A
259

MiniCPM-SALA

minicpm-sala

textinference
OOpenBMB

0.0

Agentic

25.80.00.00.00.0N/A
260

Ministral 3 (14B Reasoning 2512)

ministral-14b-latest

multimodalvisionmulti-input reasoning
Mistral AI

0.0

Agentic

35.40.00.00.00.0
241
M

Llama 3.2 11B Instruct

Meta

0.0

N/A

242
M

Llama 3.2 3B Instruct

Meta

0.0

N/A

243
M

Llama 3.2 90B Instruct

Meta

0.0

N/A

244

Page 13 of 17 · 334 models

PreviousNext

Want benchmark charts, model comparison, and pricing analytics?

Sign in to access the full interactive leaderboard with deep benchmark breakdowns and model comparison tools.

Open full leaderboard

Rankings are based on multi-dimensional evaluation across benchmark quality, inference efficiency, and cost-per-output. Scores are updated continuously and may differ from individual third-party benchmarks.

N/A
N/A
N/A
M

Llama 3.3 70B Instruct

Meta

0.0

N/A

245
N

Llama-3.3 Nemotron Super 49B v1

NVIDIA

0.0

N/A

246
M

Llama 4 Maverick

Meta

0.0

N/A

247
M

Llama 4 Scout

Meta

0.0

N/A

248

LongCat-Flash-Thinking

Meituan

0.0

N/A

249

Magistral Medium

Mistral AI

0.0

N/A

250

Magistral Small 2506

Mistral AI

0.0

N/A

251
M

MAI-Code-1-Flash

Microsoft

0.0

N/A

252
M

MAI-Thinking-1

Microsoft

0.0

N/A

253

MedGemma 4B IT

Google

0.0

N/A

254
I

Mercury 2

Inception

0.0

$0.25 in / $0.75 out

255

MiMo-V2.5

Xiaomi

0.0

$0.168 in / $0.336 out

256

MiMo-V2.5-Pro

Xiaomi

0.0

$0.435 in / $0.87 out

257

MiMo-V2-Omni

Xiaomi

0.0

N/A

258

MiMo-V2-Pro

Xiaomi

0.0

N/A

259
O

MiniCPM-SALA

OpenBMB

0.0

N/A

260

Ministral 3 (14B Reasoning 2512)

Mistral AI

0.0

N/A