R1 Distill Llama 70B Pricing
ReasoningDeepSeek · 8K tokens context
R1 Distill Llama 70B from DeepSeek costs $0.800 per 1 million input tokens and $0.800 per 1 million output tokens as of July 2026 (live OpenRouter data). The model supports a 8,192-token context window (approximately 6,144 words) with a 8K-token maximum output. A typical 1,000-token request costs $0.0008 in input charges; a 10,000-token request costs $0.0080.
| Input price | $0.800 / 1M tokens |
|---|---|
| Output price | $0.800 / 1M tokens |
| Output / input ratio | 1.0× |
| Context window | 8,192 tokens (~6,144 words) |
| Maximum output | 8,192 tokens |
| Cost per 1K tokens (input) | $0.0008 |
| Tier | Reasoning |
| Last verified |
DeepSeek R1 Distill Llama 70B is a distilled large language model based on Llama-3.3-70B-Instruct, using outputs from DeepSeek R1.
Input Price
$0.800
per 1 million tokens
Output Price
$0.800
per 1 million tokens
Context Window
8K tokens
max 8K output
Cost Examples
| Request Type | Tokens | Input Cost | Output Cost |
|---|---|---|---|
| 1,000 word article | 1,333 | $0.00107 | $0.00032 |
| 10-page document (2,500 words) | 3,333 | $0.00267 | $0.0008 |
| 1,000 lines of code | 5,000 | $0.004 | $0.0012 |
| 100K token document | 100,000 | $0.08 | $0.024 |
Output cost estimated at 30% of input token count. Use the calculator for exact figures.
Strengths
- ✓Affordable at $0.80/1M input tokens
- ✓Step-by-step chain-of-thought reasoning
- ✓Strong general-purpose performance
Limitations
- –Higher latency; reasoning tokens add to output cost
- –Smaller 8K-token context limits long-document use
Best Use Cases
Calculate R1 Distill Llama 70B Costs
Use the TokenRate calculator to convert any budget, token count, or text into exact R1 Distill Llama 70B costs — and compare across all models.
Open Calculator →R1 Distill Llama 70B — FAQ
Related Models
Related Guides