DeepSeek V4 Flash 0731 Pricing
FastDeepSeek · 1M tokens context
DeepSeek V4 Flash 0731 from DeepSeek costs $0.065 per 1 million input tokens and $0.180 per 1 million output tokens as of September 2026 (live OpenRouter data). The model supports a 1,310,720-token context window (approximately 983,040 words) with a 944K-token maximum output. A typical 1,000-token request costs $0.0001 in input charges; a 10,000-token request costs $0.0006.
| Input price | $0.065 / 1M tokens |
|---|---|
| Output price | $0.180 / 1M tokens |
| Output / input ratio | 2.8× |
| Context window | 1,310,720 tokens (~983,040 words) |
| Maximum output | 943,718 tokens |
| Cost per 1K tokens (input) | $0.0001 |
| Tier | Fast |
| Last verified |
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total.
Input Price
$0.065
per 1 million tokens
Output Price
$0.180
per 1 million tokens
Context Window
1M tokens
max 944K output
Cost Examples
| Request Type | Tokens | Input Cost | Output Cost |
|---|---|---|---|
| 1,000 word article | 1,333 | $0.0000866 | $0.000072 |
| 10-page document (2,500 words) | 3,333 | $0.000217 | $0.00018 |
| 1,000 lines of code | 5,000 | $0.000325 | $0.00027 |
| 100K token document | 100,000 | $0.0065 | $0.0054 |
Output cost estimated at 30% of input token count. Use the calculator for exact figures.
Strengths
- ✓Extremely cheap at $0.065/1M input tokens
- ✓Massive 1M-token context window
- ✓Low latency for high-throughput workloads
Limitations
- –Less capable than flagship models on complex reasoning
- –Quality and availability can vary by hosting provider
Best Use Cases
Calculate DeepSeek V4 Flash 0731 Costs
Use the TokenRate calculator to convert any budget, token count, or text into exact DeepSeek V4 Flash 0731 costs — and compare across all models.
Open Calculator →DeepSeek V4 Flash 0731 — FAQ
Related Models
Related Guides