#70Anthropic
Claude Gen-5.
Context
—
Input $/1M
—
Output $/1M
—
Modality
text
Benchmarks
How these are aggregated →Arena Elo
—
LMArena — human blind votes
AA Index
—
Artificial Analysis composite
LiveBench
—
Contamination-resistant
GPQA
—
Grad-level science reasoning
MMLU-Pro
—
Broad knowledge
SWE-bench
—
Real GitHub bugs (Verified)
Terminal-Bench
—
Agentic shell tasks
ARC-AGI
—
Abstract reasoning
Aider Polyglot
—
Multi-language code editing
Speed
—
Median output speed
Nearby on the leaderboard
- Elo 1550
- Elo 1602
- Elo 1614
- Elo 1571