#52Anthropic· released 2026-05-19

Claude 4.5 Opus.

Strong coding + long-horizon reasoning.

View on OpenRouter ↗
Context
500K
Input $/1M
$6.00
Output $/1M
$24.00
Modality
text
Arena Elo
1458
LMArena — human blind votes
AA Index
Artificial Analysis composite
LiveBench
Contamination-resistant
GPQA
81.9%
Grad-level science reasoning
MMLU-Pro
90.6%
Broad knowledge
SWE-bench
Real GitHub bugs (Verified)
Terminal-Bench
Agentic shell tasks
ARC-AGI
Abstract reasoning
Aider Polyglot
Multi-language code editing
Speed
64 tok/s
Median output speed

Profile

Claude 4.5 Opus is Anthropic's entry at position #52 on the current aggregated leaderboard. Its strongest signals are graduate-level science reasoning, very long context. Every figure below is aggregated from the public sources listed on the homepage and refreshed on each pipeline run.

Caveats worth knowing before you read too much into the position: no execution-graded coding result is public yet. As with every model on this index, treat small gaps as noise and test on your own workload before committing.

How this model's position is computed is documented in the methodology; what each benchmark actually measures is in the benchmark guide. See a figure that disagrees with its source? Report it — accepted corrections apply on the next refresh.

Nearby on the leaderboard