Gemini 2.0 Flash-Lite
The lowest-cost earlier-generation Flash model with a 1M-token context window, now shut down in favour of the Gemini 3 line.
Source: Gemini API models, Gemini API pricing, Gemini deprecations (shutdown dates), Context caching (minimums, implicit vs explicit), Metadata verification
- Text
- Image
- Text
Pricing
Every billing leg from the live catalog, in USD per 1M tokens.
| Billing leg | Rate | Batch |
|---|---|---|
| Input | $0.075 / M | $0.0375 / M |
| Output | $0.3 / M | $0.15 / M |
Cost calculator
Billing-grade math over the rate card above, including the thinking tokens most estimates miss.
Provider
Where this model runs, and what it charges.
| Platform | Input / M | Output / M | Batch | Access |
|---|---|---|---|---|
Google | $0.075 | $0.3 | ✓ | First-party API |
Price history
Has this model gotten cheaper or more expensive?
| Effective | Input / Output | Status |
|---|---|---|
| April 20, 2026 – July 10, 2026 | $0.075 / $0.3 | past |
| July 10, 2026 | $0.075 / $0.3 | current |
More from Google
Other models from the same provider.
Stop estimating. Track it.
SuperPenguin meters every Gemini 2.0 Flash-Lite call your team makes, prices it with this exact rate card, and shows spend by feature, customer, and team.
Rates are generated from SuperPenguin's live pricing catalog.