TokenRate

GLM 5.3 FlashX Pricing

Fast

Zhipu AI · 1M tokens context

GLM 5.3 FlashX from Zhipu AI costs $0.370 per 1 million input tokens and $1.25 per 1 million output tokens as of September 2026 (live OpenRouter data). The model supports a 1,048,576-token context window (approximately 786,432 words) with a 131K-token maximum output. A typical 1,000-token request costs $0.0004 in input charges; a 10,000-token request costs $0.0037.

GLM 5.3 FlashX pricing and capability summary
Input price$0.370 / 1M tokens
Output price$1.25 / 1M tokens
Output / input ratio3.4×
Context window1,048,576 tokens (~786,432 words)
Maximum output131,072 tokens
Cost per 1K tokens (input)$0.0004
TierFast
Last verified

GLM-5.3-FlashX is the high-speed variant of Z. ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s.

Live pricing from OpenRouter

Input Price

$0.370

per 1 million tokens

Output Price

$1.25

per 1 million tokens

Context Window

1M tokens

max 131K output

Cost Examples

Request TypeTokensInput CostOutput Cost
1,000 word article1,333$0.000493$0.0005
10-page document (2,500 words)3,333$0.00123$0.00125
1,000 lines of code5,000$0.00185$0.00187
100K token document100,000$0.037$0.0375

Output cost estimated at 30% of input token count. Use the calculator for exact figures.

Strengths

  • Affordable at $0.37/1M input tokens
  • Massive 1M-token context window
  • Low latency for high-throughput workloads
  • Multimodal: understands images as well as text

Limitations

  • Less capable than flagship models on complex reasoning
  • Quality and availability can vary by hosting provider

Best Use Cases

High-volume classification
Text extraction and summarization
Simple chat and Q&A
Cost-sensitive pipelines

Calculate GLM 5.3 FlashX Costs

Use the TokenRate calculator to convert any budget, token count, or text into exact GLM 5.3 FlashX costs — and compare across all models.

Open Calculator →

GLM 5.3 FlashX — FAQ

Related Models

Related Guides