Rankings

The best models, ranked.

Frontier models ranked by benchmark intelligence — reasoning, coding, math and knowledge evals — with real-world request volume shown alongside. Sort by quality to find the strongest model, or by requests to see what the market actually runs.

40 models · benchmarks 2026-07-25 · usage 2026-07-10

Top models

#ModelProviderIntelligenceCodingRequestsShare
1Claude Opus 5
anthropic/claude-opus-5
Anthropic60.778.0
2Claude Fable 5
anthropic/claude-fable-5
Anthropic59.976.52420%
3GPT-5.6 Sol
openai/gpt-5.6-sol
OpenAI58.977.4
4Kimi K3
moonshot/kimi-k3
Moonshot57.176.2
5Claude Opus 4.8
anthropic/claude-opus-4.8
Anthropic55.774.3370%
6GPT-5.6 Terra
openai/gpt-5.6-terra
OpenAI55.076.7
7GPT-5.5
openai/gpt-5.5
OpenAI54.874.935.2K3.1%
8Grok 4.5
xai/grok-4.5
xAI53.872.4
9Claude Opus 4.7
anthropic/claude-opus-4.7
Anthropic53.573.610%
10Claude Sonnet 5
anthropic/claude-sonnet-5
Anthropic53.471.5
11GPT-5.4
openai/gpt-5.4
OpenAI51.471.119.2K1.7%
12GPT-5.6 Luna
openai/gpt-5.6-luna
OpenAI51.271.4
13Glm-5.2
zai-org/glm-5.2
Zai-org51.168.816.1K1.4%
14Gemini 3.5 Flash
google/gemini-3.5-flash
Google50.270.1
15Gemini 3.6 Flash
google/gemini-3.6-flash
Google50.169.2
16Claude Sonnet 4.6
anthropic/claude-sonnet-4.6
Anthropic47.263.0
17Gemini 3.1 Pro
google/gemini-3.1-pro
Google46.568.840.7K3.5%
18MiniMax M3
minimax/m3
MiniMax44.458.6
19DeepSeek V4 Pro
deepseek/deepseek-v4-pro
Deepseek44.359.457.9K5%
20Kimi-K2.6
moonshot/kimi-k2.6
Moonshot44.261.89.4K0.8%
21Claude Opus 4.6
anthropic/claude-opus-4.6
Anthropic43.7
22Kimi-K2.7-Code
moonshot/kimi-k2.7-code
Moonshot41.960.84.3K0.4%
23Claude Opus 4.5
anthropic/claude-opus-4.5
Anthropic40.810%
24GPT-5.4 Mini
openai/gpt-5.4-mini
OpenAI40.056.160.6K5.3%
25GPT-5.4 Nano
openai/gpt-5.4-nano
OpenAI38.256.114.7K1.3%
26MiniMax M2.7
minimax/m2.7
MiniMax38.152.6
27Gemini 3 Flash
google/gemini-3-flash
Google37.8248.8K21.6%
28Grok 4.3
xai/grok-4.3
xAI37.642.234.4K3%
29GPT-5.1
openai/gpt-5.1
OpenAI36.949.4
30Gemini 3.5 Flash-Lite
google/gemini-3.5-flash-lite
Google36.549.3
31Grok 4.20 Non-Reasoning
xai/grok-4.20-0309-non-reasoning
xAI36.5
32Grok 4.20 Reasoning
xai/grok-4.20-0309-reasoning
xAI36.5
33Claude Sonnet 4.5
anthropic/claude-sonnet-4.5
Anthropic36.452.110%
34GPT-5
openai/gpt-5
OpenAI34.737.82.9K0.3%
35Qwen 3.5 397B A17B
alibaba/qwen3.5-397b-a17b
Alibaba33.748.22420%
36Grok 4
xai/grok-4
xAI33.3
37Qwen 3 Max
alibaba/qwen3-max
Alibaba31.74840%
38GPT-5 Mini
openai/gpt-5-mini
OpenAI30.934.8K3%
39O3
openai/o3
OpenAI30.41390%
40Claude Haiku 4.5
anthropic/claude-haiku-4.5
Anthropic29.643.92.9K0.2%

By provider · demand

ProviderModelsRequests
Google11766.6K
OpenAI27246.4K
Deepseek257.9K
xAI735.0K
Zai-org216.1K
Moonshot313.6K
Anthropic113.7K
Mistral12.6K
Ibm-granite12.4K
Meta51.7K
Black Forest Labs91.1K
Inworld2947
How this is built

Benchmarks for quality, usage for demand.

Intelligence and coding scores come from Artificial Analysis evals baked into the catalog on every release; request volume is real aggregate demand across the open model ecosystem, refreshed daily. For the full benchmark + pricing table, see the model leaderboard.