Qwen3.8-Max
Qwen's 2.4-trillion-parameter multimodal MoE flagship for coding, reasoning, visual understanding, and professional workflows.
Source: Alibaba Cloud Qwen3.8-Max launch, QwenCloud model capabilities, Alibaba Model Studio Qwen3.8-Max overview, OpenRouter model record, OpenRouter prompt caching
Pricing
Every billing leg from the live catalog, in USD per 1M tokens.
| Billing leg | Rate |
|---|---|
| Input | $2 / M |
| Output | $6 / M |
| Cache read | $0.25 / M |
Cost calculator
Billing-grade math over the rate card above, including the thinking tokens most estimates miss.
Hidden reasoning tokens bill as output. 1.0x = thinking off. Medium-effort reasoning typically lands near 3x; heavy reasoning runs 6x to 8x. Set it to match your workload.
Provider
Where this model runs, and what it charges.
| Platform | Input / M | Output / M | Cache read / M | Batch | Access |
|---|---|---|---|---|---|
Alibaba | $2 | $6 | $0.25 | — | via OpenRouter |
Price history
Has this model gotten cheaper or more expensive?
| Effective | Input / Output | Status |
|---|---|---|
| August 3, 2026 – August 5, 2026 | $2 / $6 | past |
| August 5, 2026 | $2 / $6 | current |
More from Qwen
Other models from the same provider.
Stop estimating. Track it.
SuperPenguin meters every Qwen3.8-Max call your team makes, prices it with this exact rate card, and shows spend by feature, customer, and team.
Rates are generated from SuperPenguin's live pricing catalog.