#51DeepSeek· released 2026-07-31

DeepSeek: DeepSeek V4 Flash 0731.

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows.

View on OpenRouter ↗
Context
1.0M
Input $/1M
$0.08
Output $/1M
$0.18
Modality
text
Arena Elo
1495
LMArena — human blind votes
AA Index
60.0
Artificial Analysis composite
LiveBench
75.0
Contamination-resistant
GPQA
80.0%
Grad-level science reasoning
MMLU-Pro
83.0%
Broad knowledge
SWE-bench
70.0%
Real GitHub bugs (Verified)
Terminal-Bench
57.0%
Agentic shell tasks
ARC-AGI
17.0%
Abstract reasoning
Aider Polyglot
81.0%
Multi-language code editing
Speed
154 tok/s
Median output speed

Profile

DeepSeek: DeepSeek V4 Flash 0731 is DeepSeek's entry at position #51 on the current aggregated leaderboard. Its strongest signals are graduate-level science reasoning, real-world software engineering, unattended agentic work, and cost efficiency. Every figure below is aggregated from the public sources listed on the homepage and refreshed on each pipeline run.

The score spread is comparatively complete for this model, but the standing advice still applies: treat gaps of a few points as ties and validate on your own workload before committing.

How this model's position is computed is documented in the methodology; what each benchmark actually measures is in the benchmark guide. See a figure that disagrees with its source? Report it — accepted corrections apply on the next refresh.

Nearby on the leaderboard