Qwen3-Coder-Next
The efficient next-generation Qwen coding MoE, available via inference gateways.
Source: Qwen (Alibaba Model Studio) pricing, Qwen models, Metadata verification, Field resolution verification
- Text
- Text
Pricing
Every billing leg from the live catalog, in USD per 1M tokens.
| Billing leg | Rate | Batch |
|---|---|---|
| Input | $0.5 / M | $0.25 / M |
| Output | $1.2 / M | $0.6 / M |
Cost calculator
Billing-grade math over the rate card above, including the thinking tokens most estimates miss.
Compare providers
The same model, priced across every platform that serves it. Lowest combined in/out rate is flagged.
| Platform | Region | Input / M | Output / M | Cache read / M | Batch | Access |
|---|---|---|---|---|---|---|
AWS Bedrock | $0.5 | $1.2 | - | ✓ | Direct | |
IonstreamLowest | - | $0.11 | $0.8 | $0.07 | - | via OpenRouter |
Parasail | - | $0.12 | $0.8 | $0.07 | - | via OpenRouter |
StreamLake | - | $0.18 | $0.9 | $0.036 | - | via OpenRouter |
Together AI | - | $0.5 | $1.2 | - | - | Direct |
Novita | - | $0.2 | $1.5 | - | - | via OpenRouter |
Alibaba | - | $0.3 | $1.5 | - | - | via OpenRouter |
Some hosts price by region. Shown rates match the selected region.
Price history
Has this model gotten cheaper or more expensive?
| Effective | Input / Output | Status |
|---|---|---|
| July 6, 2026 – July 8, 2026 | $0.5 / $1.2 | past |
| July 8, 2026 | $0.5 / $1.2 | current |
More from Qwen
Other models from the same provider.
Stop estimating. Track it.
SuperPenguin meters every Qwen3-Coder-Next call your team makes, prices it with this exact rate card, and shows spend by feature, customer, and team.
Rates are generated from SuperPenguin's live pricing catalog.