Qwen3 VL 32B Instruct Pricing
FastQwen · 131K tokens context
Qwen3 VL 32B Instruct from Qwen costs $0.104 per 1 million input tokens and $0.416 per 1 million output tokens as of July 2026 (live OpenRouter data). The model supports a 131,072-token context window (approximately 98,304 words) with a 33K-token maximum output. A typical 1,000-token request costs $0.0001 in input charges; a 10,000-token request costs $0.0010.
| Input price | $0.104 / 1M tokens |
|---|---|
| Output price | $0.416 / 1M tokens |
| Output / input ratio | 4.0× |
| Context window | 131,072 tokens (~98,304 words) |
| Maximum output | 32,768 tokens |
| Cost per 1K tokens (input) | $0.0001 |
| Tier | Fast |
| Last verified |
Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video.
Input Price
$0.104
per 1 million tokens
Output Price
$0.416
per 1 million tokens
Context Window
131K tokens
max 33K output
Cost Examples
| Request Type | Tokens | Input Cost | Output Cost |
|---|---|---|---|
| 1,000 word article | 1,333 | $0.000139 | $0.000166 |
| 10-page document (2,500 words) | 3,333 | $0.000347 | $0.000416 |
| 1,000 lines of code | 5,000 | $0.00052 | $0.000624 |
| 100K token document | 100,000 | $0.0104 | $0.0125 |
Output cost estimated at 30% of input token count. Use the calculator for exact figures.
Strengths
- ✓Extremely cheap at $0.104/1M input tokens
- ✓Low latency for high-throughput workloads
- ✓Multimodal: understands images as well as text
Limitations
- –Less capable than flagship models on complex reasoning
- –Quality and availability can vary by hosting provider
Best Use Cases
Calculate Qwen3 VL 32B Instruct Costs
Use the TokenRate calculator to convert any budget, token count, or text into exact Qwen3 VL 32B Instruct costs — and compare across all models.
Open Calculator →Qwen3 VL 32B Instruct — FAQ
Related Models
Related Guides