Qwen3.5-Flash
The Qwen3.5 fast, low-cost tier for high-volume workloads.
Source: Qwen (Alibaba Model Studio) pricing, Qwen models, Metadata verification, Metadata verification
Pricing
Every billing leg from the live catalog, in USD per 1M tokens.
| Billing leg | Rate |
|---|---|
| Input | $0.065 / M |
| Output | $0.26 / M |
Cost calculator
Billing-grade math over the rate card above, including the thinking tokens most estimates miss.
Hidden reasoning tokens bill as output. 1.0x = thinking off. Medium-effort reasoning typically lands near 3x; heavy reasoning runs 6x to 8x. Set it to match your workload.
Provider
Where this model runs, and what it charges.
| Platform | Input / M | Output / M | Batch | Access |
|---|---|---|---|---|
Alibaba | $0.065 | $0.26 | — | via OpenRouter |
Price history
Has this model gotten cheaper or more expensive?
| Effective | Input / Output | Status |
|---|---|---|
| July 10, 2026 – July 17, 2026 | $0.065 / $0.26 | past |
| July 17, 2026 – July 19, 2026 | $0.065 / $0.26 | past |
| July 19, 2026 – July 25, 2026 | $0.065 / $0.26 | past |
| July 25, 2026 | $0.065 / $0.26 | current |
More from Qwen
Other models from the same provider.
Stop estimating. Track it.
SuperPenguin meters every Qwen3.5-Flash call your team makes, prices it with this exact rate card, and shows spend by feature, customer, and team.
Rates are generated from SuperPenguin's live pricing catalog.