Rankings

The best models, ranked.

Frontier models ranked by benchmark intelligence — reasoning, coding, math and knowledge evals — with real-world request volume shown alongside. Sort by quality to find the strongest model, or by requests to see what the market actually runs.

40 models · benchmarks 2026-08-14 · usage 2026-07-10

Top models

#ModelProviderIntelligenceCodingRequestsShare
1Claude Opus 5
anthropic/claude-opus-5
Anthropic63.178.0
2Claude Fable 5
anthropic/claude-fable-5
Anthropic62.176.52420%
3GPT-5.6 Sol
openai/gpt-5.6-sol
OpenAI60.977.4
4Grok 4.6
xai/grok-4.6
xAI60.976.8
5Kimi K3
moonshot/kimi-k3
Moonshot59.776.2
6Qwen3.8 Max
alibaba/qwen3.8-max
Alibaba58.171.8
7Claude Opus 4.8
anthropic/claude-opus-4.8
Anthropic57.374.3370%
8GPT-5.6 Terra
openai/gpt-5.6-terra
OpenAI56.676.7
9GPT-5.5
openai/gpt-5.5
OpenAI56.374.935.2K3.1%
10Grok 4.5
xai/grok-4.5
xAI55.872.4
11Claude Sonnet 5
anthropic/claude-sonnet-5
Anthropic55.371.5
12Claude Opus 4.7
anthropic/claude-opus-4.7
Anthropic55.073.610%
13DeepSeek V4 Pro
deepseek/deepseek-v4-pro
Deepseek53.268.857.9K5%
14GPT-5.4
openai/gpt-5.4
OpenAI53.171.119.2K1.7%
15GLM-5.2
zai-org/glm-5.2
Zai-org52.668.816.1K1.4%
16GPT-5.6 Luna
openai/gpt-5.6-luna
OpenAI52.371.4
17Gemini 3.5 Flash
google/gemini-3.5-flash
Google52.070.1
18Gemini 3.6 Flash
google/gemini-3.6-flash
Google51.669.2
19Claude Sonnet 4.6
anthropic/claude-sonnet-4.6
Anthropic48.463.0
20Gemini 3.1 Pro
google/gemini-3.1-pro
Google47.768.840.7K3.5%
21Qwen3.7 Max
alibaba/qwen3.7-max
Alibaba46.766.0
22MiniMax M3
minimax/m3
MiniMax45.458.6
23Kimi-K2.6
moonshot/kimi-k2.6
Moonshot45.161.89.4K0.8%
24Claude Opus 4.6
anthropic/claude-opus-4.6
Anthropic44.9
25Kimi-K2.7-Code
moonshot/kimi-k2.7-code
Moonshot43.060.84.3K0.4%
26Claude Opus 4.5
anthropic/claude-opus-4.5
Anthropic41.910%
27GPT-5.4 Mini
openai/gpt-5.4-mini
OpenAI40.956.160.6K5.3%
28GPT-5.4 Nano
openai/gpt-5.4-nano
OpenAI39.756.114.7K1.3%
29Qwen3.7 Plus
alibaba/qwen3.7-plus
Alibaba39.455.9
30MiniMax M2.7
minimax/m2.7
MiniMax38.952.6
31Gemini 3 Flash
google/gemini-3-flash
Google38.7248.8K21.6%
32Grok 4.3
xai/grok-4.3
xAI37.942.234.4K3%
33GPT-5.1
openai/gpt-5.1
OpenAI37.549.4
34Claude Sonnet 4.5
anthropic/claude-sonnet-4.5
Anthropic37.452.110%
35Gemini 3.5 Flash-Lite
google/gemini-3.5-flash-lite
Google37.449.3
36Grok 4.20 Non-Reasoning
xai/grok-4.20-0309-non-reasoning
xAI37.4
37Grok 4.20 Reasoning
xai/grok-4.20-0309-reasoning
xAI37.4
38GPT-5
openai/gpt-5
OpenAI35.337.82.9K0.3%
39Qwen 3.5 397B A17B
alibaba/qwen3.5-397b-a17b
Alibaba34.348.22420%
40Grok 4
xai/grok-4
xAI34.1

By provider · demand

ProviderModelsRequests
Google11766.6K
OpenAI27246.4K
Deepseek257.9K
xAI835.0K
Zai-org216.1K
Moonshot313.6K
Anthropic113.7K
Mistral12.6K
Ibm-granite12.4K
Meta51.7K
Black Forest Labs91.1K
Inworld2947
How this is built

Benchmarks for quality, usage for demand.

Intelligence and coding scores come from Artificial Analysis evals baked into the catalog on every release; request volume is real aggregate demand across the open model ecosystem, refreshed daily. For the full benchmark + pricing table, see the model leaderboard.