Rankings

The best models, ranked.

Frontier models ranked by benchmark intelligence — reasoning, coding, math and knowledge evals — with real-world request volume shown alongside. Sort by quality to find the strongest model, or by requests to see what the market actually runs.

40 models · benchmarks 2026-09-24 · usage 2026-07-10

Top models

#ModelProviderIntelligenceCodingRequestsShare
1Claude Opus 5.5
anthropic/claude-opus-5.5
Anthropic57.6
2Claude Fable 5.1
anthropic/claude-fable-5.1
Anthropic53.481.6
3GPT-6 Astra
openai/gpt-6-astra
OpenAI52.776.9
4Claude Opus 5
anthropic/claude-opus-5
Anthropic50.878.0
5Claude Fable 5
anthropic/claude-fable-5
Anthropic49.676.52420%
6GPT-6 Sol
openai/gpt-6-sol
OpenAI47.5
7GPT-5.6 Sol
openai/gpt-5.6-sol
OpenAI47.077.4
8Grok 4.7
xai/grok-4.7
xAI46.4
9Qwen3.8 Max
alibaba/qwen3.8-max
Alibaba45.476.2
10Glm-5.3
zai-org/glm-5.3
Zai-org44.874.8
11Grok 4.6
xai/grok-4.6
xAI44.376.8
12Kimi K3
moonshot/kimi-k3
Moonshot43.676.2
13GPT-5.6 Terra
openai/gpt-5.6-terra
OpenAI42.176.7
14Claude Opus 4.8
anthropic/claude-opus-4.8
Anthropic41.874.3370%
15Glm-5.3-Flash
zai-org/glm-5.3-flash
Zai-org41.871.5
16Gemini 3.8 Flash
google/gemini-3.8-flash
Google40.976.3
17Claude Opus 4.7
anthropic/claude-opus-4.7
Anthropic40.773.610%
18Gemini 3.7 Flash
google/gemini-3.7-flash
Google39.671.5
19GPT-5.4
openai/gpt-5.4
OpenAI39.071.119.2K1.7%
20Grok 4.5
xai/grok-4.5
xAI38.872.4
21GPT-5.5
openai/gpt-5.5
OpenAI38.474.935.2K3.1%
22Claude Sonnet 5
anthropic/claude-sonnet-5
Anthropic38.271.5
23GPT-5.6 Luna
openai/gpt-5.6-luna
OpenAI37.371.4
24GPT-6 Luna
openai/gpt-6-luna
OpenAI37.3
25DeepSeek V4 Pro
deepseek/deepseek-v4-pro
Deepseek36.068.857.9K5%
26Gemini 3.6 Flash
google/gemini-3.6-flash
Google34.069.2
27GLM-5.2
zai-org/glm-5.2
Zai-org33.768.816.1K1.4%
28Qwen3.8-27b
qwen/qwen3.8-27b
Alibaba Qwen33.768.1
29Gemini 3.5 Flash
google/gemini-3.5-flash
Google33.6
30Claude Opus 4.6
anthropic/claude-opus-4.6
Anthropic31.9
31Claude Sonnet 4.6
anthropic/claude-sonnet-4.6
Anthropic30.163.0
32Gemini 3.1 Pro
google/gemini-3.1-pro
Google29.768.840.7K3.5%
33Qwen3.7 Max
alibaba/qwen3.7-max
Alibaba29.566.0
34MiniMax M3
minimax/m3
MiniMax29.258.6
35Claude Opus 4.5
anthropic/claude-opus-4.5
Anthropic29.110%
36Kimi-K2.6
moonshot/kimi-k2.6
Moonshot27.061.89.4K0.8%
37Gemini 3 Flash
google/gemini-3-flash
Google26.3248.8K21.6%
38Kimi-K2.7-Code
moonshot/kimi-k2.7-code
Moonshot25.860.84.3K0.4%
39Qwen3.7 Plus
alibaba/qwen3.7-plus
Alibaba25.255.9
40Grok 4.20 Non-Reasoning
xai/grok-4.20-0309-non-reasoning
xAI25.2

By provider · demand

ProviderModelsRequests
Google13766.6K
OpenAI30246.4K
Deepseek257.9K
xAI735.0K
Zai-org416.1K
Moonshot313.6K
Anthropic133.7K
Mistral12.6K
Ibm-granite12.4K
Meta51.7K
Black Forest Labs91.1K
Inworld2947
How this is built

Benchmarks for quality, usage for demand.

Intelligence and coding scores come from Artificial Analysis evals baked into the catalog on every release; request volume is real aggregate demand across the open model ecosystem, refreshed daily. For the full benchmark + pricing table, see the model leaderboard.