TokenRate

Llama 4 Maverick Pricing

Balanced

Meta · 1M tokens context

Llama 4 Maverick from Meta costs $0.200 per 1 million input tokens and $0.800 per 1 million output tokens as of July 2026 (live OpenRouter data). The model supports a 1,048,576-token context window (approximately 786,432 words) with a 16K-token maximum output. A typical 1,000-token request costs $0.0002 in input charges; a 10,000-token request costs $0.0020.

Llama 4 Maverick pricing and capability summary
Input price$0.200 / 1M tokens
Output price$0.800 / 1M tokens
Output / input ratio4.0×
Context window1,048,576 tokens (~786,432 words)
Maximum output16,384 tokens
Cost per 1K tokens (input)$0.0002
TierBalanced
Last verified

Llama 4 Maverick is Meta's flagship open-weight model in the Llama 4 generation — multimodal, 1M context, and competitive with GPT-4o at a fraction of the API cost.

Live pricing from OpenRouter

Input Price

$0.200

per 1 million tokens

Output Price

$0.800

per 1 million tokens

Context Window

1M tokens

max 16K output

Cost Examples

Request TypeTokensInput CostOutput Cost
1,000 word article1,333$0.000267$0.00032
10-page document (2,500 words)3,333$0.000667$0.0008
1,000 lines of code5,000$0.001$0.0012
100K token document100,000$0.02$0.024

Output cost estimated at 30% of input token count. Use the calculator for exact figures.

Strengths

  • Frontier-quality open-weight model
  • Native multimodal
  • 1M context at $0.40/1M
  • Self-hostable

Limitations

  • Hosting large MoE at scale requires careful infrastructure planning

Best Use Cases

Production multimodal apps
On-prem frontier replacement
Long-context chat

Calculate Llama 4 Maverick Costs

Use the TokenRate calculator to convert any budget, token count, or text into exact Llama 4 Maverick costs — and compare across all models.

Open Calculator →

Llama 4 Maverick — FAQ

Related Models

Related Guides