Kimi K3
Moonshot's July 2026 frontier Kimi K3: a 2.8T-parameter MoE with native vision, a 1M-token context window, and always-on max reasoning.
Source: Moonshot AI pricing, Kimi models
Pricing
Every billing leg from the live catalog, in USD per 1M tokens.
| Billing leg | Rate |
|---|---|
| Input | $3 / M |
| Output | $15 / M |
| Cache read | $0.3 / M |
Cost calculator
Billing-grade math over the rate card above, including the thinking tokens most estimates miss.
Hidden reasoning tokens bill as output. 1.0x = thinking off. Medium-effort reasoning typically lands near 3x; heavy reasoning runs 6x to 8x. Set it to match your workload.
Provider
Where this model runs, and what it charges.
| Platform | Input / M | Output / M | Cache read / M | Batch | Access |
|---|---|---|---|---|---|
Moonshot AI | $3 | $15 | $0.3 | — | via OpenRouter |
Price history
Has this model gotten cheaper or more expensive?
More from Moonshot AI (Kimi)
Other models from the same provider.
Stop estimating. Track it.
SuperPenguin meters every Kimi K3 call your team makes, prices it with this exact rate card, and shows spend by feature, customer, and team.
Rates are generated from SuperPenguin's live pricing catalog.