This model is not part of the weekly automatic sync. The figures below were confirmed on 2026-08-13.
1 million token long context + inference support high-speed MoE model
DeepSeek-V4 Flash — DeepSeek. Context window 1,000,000 tokens. API pricing (per 1M tokens): Input $0.14, Output $0.28. License: Open-Weight. LMArena 1,432. GPQA Diamond 91%.
| Provider | DeepSeek |
|---|---|
| Tier | Mid |
| Context window | 1,000,000 tokens |
| Max output | 8,192 tokens |
| License | Open-Weight |
| API pricing (per 1M tokens) · Input | $0.14 |
| API pricing (per 1M tokens) · Output | $0.28 |
| LMArena | 1,432 (deepseek-v4-flash) |
| GPQA Diamond | 91% |
LMArena (CC-BY 4.0, best reasoning setting, 2026-10-02) · GPQA Diamond: Epoch AI (CC-BY 4.0)
Compare in the full catalog