GLM 5.3 FlashX pricing
$$ Standard$0.37/M input · $1.25/M output. List prices from provider data; in idapt, chat is metered per token with transparent, usage-based pricing at the prices shown in the app.
Full rate table
| Rate | Price | Unit |
|---|---|---|
| Input | 0.37 | /M tokens |
| Output | 1.25 | /M tokens |
| Cache read | 0.09 | /M tokens |
What typical work costs
| Workload | Tokens | List price |
|---|---|---|
| Quick question | 1K tokens in, 500 out | $0.0010 |
| Long-document summary | 60K tokens in, 2K out | $0.02 |
| Agent working session | 400K tokens in, 40K out across a day | $0.20 |
Price your own workload
- Per call
- $0.0010
- Per month
- $0.10
Compare up to four models at once in the cost calculator.
Frequently asked
What does GLM 5.3 FlashX cost per 1M tokens?
GLM 5.3 FlashX lists at $0.37 per 1M input tokens and $1.25 per 1M output tokens.
Does GLM 5.3 FlashX discount cached input tokens?
Yes. Cached input tokens are billed at $0.09 per 1M, versus $0.37 for fresh input. Prompts that repeat a long prefix (a system prompt, attached files) benefit the most.
What does a typical chat message with GLM 5.3 FlashX cost?
A message with about 1,000 input tokens and a 500-token reply costs roughly $0.0010 at list price. Long documents and long replies scale that linearly with token counts.
Is GLM 5.3 FlashX cheaper than GLM 5?
GLM 5.3 FlashX has the lower blended price at a 3:1 input:output mix. GLM 5.3 FlashX: $0.37 in / $1.25 out per 1M. GLM 5: $0.60 in / $1.92 out per 1M.
How does idapt bill GLM 5.3 FlashX?
Chat with GLM 5.3 FlashX in idapt is metered per token with transparent, usage-based pricing at the prices shown in the app. Subscriptions include a weekly usage allowance before any balance is touched.
Part of the GLM 5.3 FlashX model page · cheapest models board · price changes