TokenRate

GLM 5.3 Prime Pricing

Balanced

Zhipu AI · 1M tokens context

GLM 5.3 Prime from Zhipu AI costs $2.80 per 1 million input tokens and $8.80 per 1 million output tokens as of September 2026 (live OpenRouter data). The model supports a 1,000,000-token context window (approximately 750,000 words) with a 131K-token maximum output. A typical 1,000-token request costs $0.0028 in input charges; a 10,000-token request costs $0.0280.

GLM 5.3 Prime pricing and capability summary
Input price$2.80 / 1M tokens
Output price$8.80 / 1M tokens
Output / input ratio3.1×
Context window1,000,000 tokens (~750,000 words)
Maximum output131,072 tokens
Cost per 1K tokens (input)$0.0028
TierBalanced
Last verified

GLM-5.3-Prime is the high-speed variant of Z. ai's GLM-5.3, inheriting its full capabilities while delivering 1.5–2× the output throughput through inference acceleration.

Live pricing from OpenRouter

Input Price

$2.80

per 1 million tokens

Output Price

$8.80

per 1 million tokens

Context Window

1M tokens

max 131K output

Cost Examples

Request TypeTokensInput CostOutput Cost
1,000 word article1,333$0.00373$0.00352
10-page document (2,500 words)3,333$0.00933$0.0088
1,000 lines of code5,000$0.014$0.0132
100K token document100,000$0.28$0.264

Output cost estimated at 30% of input token count. Use the calculator for exact figures.

Strengths

  • Massive 1M-token context window
  • Strong general-purpose performance
  • Strong general-purpose performance

Limitations

  • Quality and availability can vary by hosting provider
  • Quality and availability can vary by hosting provider

Best Use Cases

Customer-facing AI apps
Code generation and review
Content creation at scale
Conversational interfaces

Calculate GLM 5.3 Prime Costs

Use the TokenRate calculator to convert any budget, token count, or text into exact GLM 5.3 Prime costs — and compare across all models.

Open Calculator →

GLM 5.3 Prime — FAQ

Related Models

Related Guides