GLM 4.6 pricing
$$ Standard$0.50/M input · $2.00/M output. List prices from provider data; in idapt, chat is metered per token with transparent, usage-based pricing at the prices shown in the app.
Full rate table
| Rate | Price | Unit |
|---|---|---|
| Input | 0.5 | /M tokens |
| Output | 2 | /M tokens |
| Cache read | 0.09999999999999999 | /M tokens |
What typical work costs
| Workload | Tokens | List price |
|---|---|---|
| Quick question | 1K tokens in, 500 out | $0.0015 |
| Long-document summary | 60K tokens in, 2K out | $0.03 |
| Agent working session | 400K tokens in, 40K out across a day | $0.28 |
Price your own workload
- Per call
- $0.0015
- Per month
- $0.15
Compare up to four models at once in the cost calculator.
Frequently asked
What does GLM 4.6 cost per 1M tokens?
GLM 4.6 lists at $0.50 per 1M input tokens and $2.00 per 1M output tokens.
Does GLM 4.6 discount cached input tokens?
Yes. Cached input tokens are billed at $0.10 per 1M, versus $0.50 for fresh input. Prompts that repeat a long prefix (a system prompt, attached files) benefit the most.
What does a typical chat message with GLM 4.6 cost?
A message with about 1,000 input tokens and a 500-token reply costs roughly $0.0015 at list price. Long documents and long replies scale that linearly with token counts.
Is GLM 4.6 cheaper than GLM 4.5V?
GLM 4.6 has the lower blended price at a 3:1 input:output mix. GLM 4.6: $0.50 in / $2.00 out per 1M. GLM 4.5V: $0.60 in / $1.80 out per 1M.
How does idapt bill GLM 4.6?
Chat with GLM 4.6 in idapt is metered per token with transparent, usage-based pricing at the prices shown in the app. Subscriptions include a weekly usage allowance before any balance is touched.
Part of the GLM 4.6 model page · cheapest models board · price changes