Skip to content
— LLM Leaderboard —

Ranked by the people
who use them.

Community-driven rankings of the leading large language models, scored by accredited AI practitioners across the capabilities that matter in real work: reasoning, engineering and conversation. No vendor benchmarks, no marketing claims. Just the verdict of the people who build with them every day.

— Overall standings —

The models out
in front.

1
AnthropicAnthropic

Claude Fable 5

98/100
Overall
2
AnthropicAnthropic

Claude Opus 5

96/100
Overall
3
OpenAIOpenAI

GPT-5.6 Sol

94/100
Overall
— Capability by capability —

Where each model
actually wins.

The top five performers in every category, scored on the work practitioners do most.

Professional Reasoning

Complex problem-solving and analytical judgement in real-world scenarios.

  • Claude Fable 598
  • Claude Opus 596
  • GPT-5.6 Sol94
  • Grok 4.693
  • Claude Opus 4.891

Software Engineering

Code generation, debugging and algorithmic thinking under real constraints.

  • Claude Fable 599
  • Claude Opus 596
  • GPT-5.6 Sol96
  • Kimi K394
  • Grok 4.693

Conversational Intelligence

Reading intent, tone and nuance across natural language exchanges.

  • Claude Opus 597
  • Claude Fable 596
  • Claude Opus 4.894
  • Gemini 3.1 Pro92
  • GPT-5.6 Sol92
— Top ranked —

The top-rated
models.

The highest-rated models, side by side. Tap any score column to re-rank the board by that capability.

#Model
1
Anthropic
Claude Fable 5
Anthropic
98999698
2
Anthropic
Claude Opus 5
Anthropic
96969796
3
OpenAI
GPT-5.6 Sol
OpenAI
94969294
4
Anthropic
Claude Opus 4.8
Anthropic
91929492
5
xAI
Grok 4.6
xAI
93938992
6
Moonshot AI
Kimi K3
Moonshot AI
91948991
7
Google
Gemini 3.1 Pro
Google
89859289
8
Anthropic
Claude Sonnet 5
Anthropic
87889088
9
Qwen
Qwen3.8 Max
Qwen
88878587
10
OpenAI
GPT-5.6 Terra
OpenAI
86868786
11
OpenAI
GPT-5.5
OpenAI
85848685
12
xAI
Grok 4.5
xAI
85848585
13
Anthropic
Claude Opus 4.7
Anthropic
84848484
14
Meta
Muse Spark 1.2
Meta
84838584
15
DeepSeek
DeepSeek V4 Pro
DeepSeek
83858083
16
Google
Gemini 3.7 Flash
Google
82848483
17
Z.ai
GLM-5.2
Z.ai
81827981
18
Google
Gemini 3.6 Flash
Google
79808280
19
OpenAI
GPT-5.6 Luna
OpenAI
77778279
20
DeepSeek
DeepSeek V4 Flash
DeepSeek
78807778
Help shape the rankings

Scored by the people
who build with AI.

Our leaderboard is ranked by accredited practitioners. Join the Institute to lend your expertise to the standard.