GPT-6 Astra
OpenAI's most capable flagship for complex reasoning, coding, computer use, research, and document creation. 1.05M context with tiered pricing above 272K input tokens; reasoning.effort supports low through max.
Source: OpenAI GPT-6 Astra model page, OpenAI model guidance, OpenAI pricing, Prompt caching (1,024 minimum, 30m TTL), ChatGPT release notes (2026-09-03)
- Text
- Image
- Text
Pricing
Every billing leg from the live catalog, in USD per 1M tokens.
| Billing leg | Standard | Long context | Batch |
|---|---|---|---|
| Input | $10 / M | $20 / M | $5 / M |
| Output | $50 / M | $75 / M | $25 / M |
| Cache read | $1 / M | $2 / M | $0.5 / M |
| Cache write | $12.5 / M | $25 / M | $6.25 / M |
Long-context rates apply above 272,000 input tokens.
Cost calculator
Billing-grade math over the rate card above, including the thinking tokens most estimates miss.
Hidden reasoning tokens bill as output. 1.0x = thinking off. Medium-effort reasoning typically lands near 3x; heavy reasoning runs 6x to 8x. Set it to match your workload.
Provider
Where this model runs, and what it charges.
| Platform | Input / M | Output / M | Cache read / M | Batch | Access |
|---|---|---|---|---|---|
OpenAI | $10 | $50 | $1 | ✓ | First-party API |
Price history
Has this model gotten cheaper or more expensive?
More from OpenAI
Other models from the same provider.
Stop estimating. Track it.
SuperPenguin meters every GPT-6 Astra call your team makes, prices it with this exact rate card, and shows spend by feature, customer, and team.
Rates are generated from SuperPenguin's live pricing catalog.