TokenRate

GLM 4.7 Flash Pricing

Fast

Zhipu AI · 200K tokens context

GLM 4.7 Flash from Zhipu AI costs $0.061 per 1 million input tokens and $0.400 per 1 million output tokens as of September 2026 (live OpenRouter data). The model supports a 200,000-token context window (approximately 150,000 words) with a 118K-token maximum output. A typical 1,000-token request costs $0.0001 in input charges; a 10,000-token request costs $0.0006.

GLM 4.7 Flash pricing and capability summary
Input price$0.061 / 1M tokens
Output price$0.400 / 1M tokens
Output / input ratio6.6×
Context window200,000 tokens (~150,000 words)
Maximum output117,964 tokens
Cost per 1K tokens (input)$0.0001
TierFast
Last verified

As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency.

Live pricing from OpenRouter

Input Price

$0.061

per 1 million tokens

Output Price

$0.400

per 1 million tokens

Context Window

200K tokens

max 118K output

Cost Examples

Request TypeTokensInput CostOutput Cost
1,000 word article1,333$0.0000806$0.00016
10-page document (2,500 words)3,333$0.000202$0.0004
1,000 lines of code5,000$0.000302$0.0006
100K token document100,000$0.00605$0.012

Output cost estimated at 30% of input token count. Use the calculator for exact figures.

Strengths

  • ✓Extremely cheap at $0.061/1M input tokens
  • ✓Large 200K-token context window
  • ✓Low latency for high-throughput workloads

Limitations

  • –Less capable than flagship models on complex reasoning
  • –Quality and availability can vary by hosting provider

Best Use Cases

High-volume classification
Text extraction and summarization
Simple chat and Q&A
Cost-sensitive pipelines

Calculate GLM 4.7 Flash Costs

Use the TokenRate calculator to convert any budget, token count, or text into exact GLM 4.7 Flash costs — and compare across all models.

Open Calculator →

GLM 4.7 Flash — FAQ

Related Models

Related Guides