TokenRate

DeepSeek V4 Flash 0731 Pricing

Fast

DeepSeek · 1M tokens context

DeepSeek V4 Flash 0731 from DeepSeek costs $0.065 per 1 million input tokens and $0.180 per 1 million output tokens as of September 2026 (live OpenRouter data). The model supports a 1,310,720-token context window (approximately 983,040 words) with a 944K-token maximum output. A typical 1,000-token request costs $0.0001 in input charges; a 10,000-token request costs $0.0006.

DeepSeek V4 Flash 0731 pricing and capability summary
Input price$0.065 / 1M tokens
Output price$0.180 / 1M tokens
Output / input ratio2.8×
Context window1,310,720 tokens (~983,040 words)
Maximum output943,718 tokens
Cost per 1K tokens (input)$0.0001
TierFast
Last verified

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total.

Live pricing from OpenRouter

Input Price

$0.065

per 1 million tokens

Output Price

$0.180

per 1 million tokens

Context Window

1M tokens

max 944K output

Cost Examples

Request TypeTokensInput CostOutput Cost
1,000 word article1,333$0.0000866$0.000072
10-page document (2,500 words)3,333$0.000217$0.00018
1,000 lines of code5,000$0.000325$0.00027
100K token document100,000$0.0065$0.0054

Output cost estimated at 30% of input token count. Use the calculator for exact figures.

Strengths

  • Extremely cheap at $0.065/1M input tokens
  • Massive 1M-token context window
  • Low latency for high-throughput workloads

Limitations

  • Less capable than flagship models on complex reasoning
  • Quality and availability can vary by hosting provider

Best Use Cases

High-volume classification
Text extraction and summarization
Simple chat and Q&A
Cost-sensitive pipelines

Calculate DeepSeek V4 Flash 0731 Costs

Use the TokenRate calculator to convert any budget, token count, or text into exact DeepSeek V4 Flash 0731 costs — and compare across all models.

Open Calculator →

DeepSeek V4 Flash 0731 — FAQ

Related Models

Related Guides