Gemini 2.0 Flash-Lite
The lowest-cost earlier-generation Flash model with a 1M-token context window, now shut down in favour of the Gemini 3 line.
Source: Gemini API models, Gemini API pricing, Gemini deprecations (shutdown dates), Context caching (minimums, implicit vs explicit), Metadata verification
Pricing
Every billing leg from the live catalog, in USD per 1M tokens.
| Billing leg | Rate | Batch |
|---|---|---|
| Input | $0.075 / M | $0.0375 / M |
| Output | $0.3 / M | $0.15 / M |
Cost calculator
Billing-grade math over the rate card above, including the thinking tokens most estimates miss.
Provider
Where this model runs, and what it charges.
| Platform | Input / M | Output / M | Batch | Access |
|---|---|---|---|---|
Google | $0.075 | $0.3 | ✓ | First-party API |
Price history
Has this model gotten cheaper or more expensive?
| Effective | Input / Output | Status |
|---|---|---|
| April 20, 2026 – July 10, 2026 | $0.075 / $0.3 | past |
| July 10, 2026 | $0.075 / $0.3 | current |
More from Google
Other models from the same provider.
Stop estimating. Track it.
SuperPenguin meters every Gemini 2.0 Flash-Lite call your team makes, prices it with this exact rate card, and shows spend by feature, customer, and team.
Rates are generated from SuperPenguin's live pricing catalog.