#04Moonshot· released 2026-07-16

MoonshotAI: Kimi K3.

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

View on OpenRouter ↗
Context
1.0M
Input $/1M
$3.00
Output $/1M
$15.00
Modality
text · image · video
Arena Elo
1607
LMArena — human blind votes
AA Index
72.0
Artificial Analysis composite
LiveBench
83.9
Contamination-resistant
GPQA
89.4%
Grad-level science reasoning
MMLU-Pro
88.8%
Broad knowledge
SWE-bench
78.1%
Real GitHub bugs (Verified)
Terminal-Bench
62.5%
Agentic shell tasks
ARC-AGI
23.9%
Abstract reasoning
Aider Polyglot
84.8%
Multi-language code editing
Speed
46 tok/s
Median output speed

Profile

MoonshotAI: Kimi K3 is Moonshot's entry at position #4 on the current aggregated leaderboard. Its strongest signals are human preference voting, graduate-level science reasoning, real-world software engineering, and unattended agentic work. Every figure below is aggregated from the public sources listed on the homepage and refreshed on each pipeline run.

The score spread is comparatively complete for this model, but the standing advice still applies: treat gaps of a few points as ties and validate on your own workload before committing.

How this model's position is computed is documented in the methodology; what each benchmark actually measures is in the benchmark guide. See a figure that disagrees with its source? Report it — accepted corrections apply on the next refresh.

Nearby on the leaderboard