Conversation / Quote
GLM-5.3-FlashX API Price
Zhipu BigModel · glm-5.3-flashx
| Billing | CNY · 1M tokens |
|---|---|
| Context Window | 1000K tokens |
| Max Output | 128K tokens |
| Modalities | Text / Image / Video |
| Deep Thinking | Supported |
| Verified Date | 2026-09-19 |
| Official pricing page | docs.bigmodel.cn/cn/guide/start/pricing.md ↗ |
Source text: the note below is the model's official text, kept in its original Chinese and not translated.
极速推理版(峰值输出最高 200 tokens/s,约为 Flash 的 5 倍);官方 Model Code glm-5.3-flashx;强制思考模式(thinking.type 仅支持 enabled,不可关闭);缓存存储限时免费;价格约为 Flash 的 2.5 倍
Report a pricing issue
Found inaccurate pricing or model information? We will review it.
Monthly Cost Estimate
30 days · no cache hits · list price
Free quotas, tiered pricing, batch discounts, and time-based pricing are not included; refer to official bills.
Zhipu BigModel Related Models
FAQ
How much does the GLM-5.3-FlashX API cost?
Input ¥2/1M tokens, output ¥7/1M tokens
How large is the GLM-5.3-FlashX context window?
1000K tokens, up to 128K tokens of output. Handles long documents, multi-turn conversations and large-scale inputs.
Which input modalities does GLM-5.3-FlashX support?
Text / Image / Video. Supports deep thinking / reasoning.
What is the approximate monthly cost of GLM-5.3-FlashX?
Depends on volume: light use (1M input + 300K output per day) ≈ ¥123/month; medium (5M + 1.5M) ≈ ¥615/month; heavy agents (10M + 3M) ≈ ¥1,230/month. Cache hits, batch discounts and tiered pricing can lower this further.
How does GLM-5.3-FlashX compare with peers on value?
Zhipu BigModel GLM-5.3-FlashX is priced low among comparable models; compare against similar models. See the Pricemon market table for a full comparison.
What use cases suit GLM-5.3-FlashX?
Suited to applications in its domain.The long context supports deep-analysis scenarios.
Does GLM-5.3-FlashX have a free tier?
Check Zhipu BigModel's official pricing page for the latest free quota, trial policy and tiered pricing. Policies differ by vendor and change over time.