Qwen3.8 Max (0902) pricing
$$$ Premium$2.00/M input · $6.00/M output. List prices from provider data; in idapt, chat is metered per token with transparent, usage-based pricing at the prices shown in the app.
Full rate table
| Rate | Price | Unit |
|---|---|---|
| Input | 2 | /M tokens |
| Output | 6 | /M tokens |
| Cache read | 0.25 | /M tokens |
| Cache write | 2.5 | /M tokens |
What typical work costs
| Workload | Tokens | List price |
|---|---|---|
| Quick question | 1K tokens in, 500 out | $0.0050 |
| Long-document summary | 60K tokens in, 2K out | $0.13 |
| Agent working session | 400K tokens in, 40K out across a day | $1.04 |
Price your own workload
- Per call
- $0.0050
- Per month
- $0.50
Compare up to four models at once in the cost calculator.
Frequently asked
What does Qwen3.8 Max (0902) cost per 1M tokens?
Qwen3.8 Max (0902) lists at $2.00 per 1M input tokens and $6.00 per 1M output tokens.
Does Qwen3.8 Max (0902) discount cached input tokens?
Yes. Cached input tokens are billed at $0.25 per 1M, versus $2.00 for fresh input. Prompts that repeat a long prefix (a system prompt, attached files) benefit the most.
What does a typical chat message with Qwen3.8 Max (0902) cost?
A message with about 1,000 input tokens and a 500-token reply costs roughly $0.0050 at list price. Long documents and long replies scale that linearly with token counts.
Is Qwen3.8 Max (0902) cheaper than Qwen3.8 2.4T A95B?
Qwen3.8 Max (0902) has the lower blended price at a 3:1 input:output mix. Qwen3.8 Max (0902): $2.00 in / $6.00 out per 1M. Qwen3.8 2.4T A95B: $2.00 in / $6.00 out per 1M.
How does idapt bill Qwen3.8 Max (0902)?
Chat with Qwen3.8 Max (0902) in idapt is metered per token with transparent, usage-based pricing at the prices shown in the app. Subscriptions include a weekly usage allowance before any balance is touched.
Part of the Qwen3.8 Max (0902) model page · cheapest models board · price changes