Gemini 3.8 Flash
1,048,576 context · 65,536 max output
API ID
- Stable
Model ID reference
Copy the exact Gemini API model code, from gemini-3.8-flash to preview models, with input and output limits, status and Google’s shutdown dates. Free, and it runs in your browser.
16 models. Click an ID to copy it.
1,048,576 context · 65,536 max output
API ID
1,048,576 context · 65,536 max output
API ID
1,048,576 context · 65,536 max output
API ID
1,048,576 context · 65,536 max output
API ID
Replacement: gemini-3.8-flash
Google calls it its previous-generation Flash model.
API ID
Requests to gemini-3.7-flash are automatically routed to gemini-3.8-flash.
API ID
Requests to gemini-3.5-flash are automatically routed to gemini-3.6-flash, although Google’s quickstart still uses this code in its examples.
1,048,576 context · 65,536 max output
API ID
Replacement: gemini-3.6-flash
No shutdown date announced; Google recommends gemini-3.6-flash.
1,048,576 context · 65,536 max output
API ID
Not deprecated, but Google limits the 2.5 models to accounts that have used them before. New projects should use Gemini 3.8 Flash or 3.5 Flash-Lite.
1,048,576 context · 65,536 max output
API ID
Not deprecated, but Google limits the 2.5 models to accounts that have used them before. New projects should use Gemini 3.8 Flash or 3.5 Flash-Lite.
1,048,576 context · 65,536 max output
API ID
Not deprecated, but Google limits the 2.5 models to accounts that have used them before. New projects should use Gemini 3.8 Flash or 3.5 Flash-Lite.
1,048,576 context · 65,536 max output
API ID
Shuts down 7 May 2027 at the earliest. Replacement: gemini-3.5-flash-lite
API ID
Retired 1 Jun 2026. Replacement: gemini-3.6-flash
API ID
Retired 1 Jun 2026. Replacement: gemini-3.1-flash-lite
API ID
Retired 25 May 2026. Replacement: gemini-3.1-flash-lite
API ID
Retired 31 Mar 2026. Replacement: gemini-3.1-flash-lite
API ID
Retired 9 Mar 2026. Replacement: gemini-3.1-pro-preview
Steps
Method
Google calls the model ID a “model code”: the string you pass as model to the Gemini API, such as gemini-3.8-flash. This page lists the text models on the Gemini Developer API (Google AI Studio keys) from Google’s models and deprecations pages. Image, audio, music, video and embedding models are left out.
Google sorts every code into one of four version patterns:
gemini-3.8-flash.gemini-3.1-pro-preview, may be used in production, typically have billing enabled, may have tighter rate limits and can be deprecated with at least two weeks’ notice.gemini-flash-latest, are hot-swapped with every new release of a model variation. Google emails two weeks’ notice before a breaking change behind one.Google notes this naming scheme dates from September 2025, so older models may follow different patterns: use the exact code from the list.
For new projects Google recommends Gemini 3.8 Flash (gemini-3.8-flash) or Gemini 3.5 Flash-Lite (gemini-3.5-flash-lite). Both take up to 1,048,576 input tokens and write up to 65,536 output tokens. The Gemini 2.5 models are not deprecated, but Google now limits them to accounts that have used them before, so a new key may not be able to call gemini-2.5-pro.
Two recent codes are kept alive by redirection: requests to gemini-3.7-flash go to gemini-3.8-flash, and requests to gemini-3.5-flash go to gemini-3.6-flash. Your code keeps working, but you get a different model from the one you named, with its own behaviour. Google’s quickstart still uses gemini-3.5-flash in its examples, so copy-pasted code may already be redirected.
Google’s deprecations page lists the earliest date each model might be shut down and says the exact date is confirmed with advance notice. Gemini 3.1 Flash-Lite, for example, can shut down from 7 May 2027 and the replacement is gemini-3.5-flash-lite. Once a model is shut down its endpoint stops responding: gemini-2.0-flash was switched off on 1 Jun 2026, with gemini-3.6-flash as the replacement.
Gemini models are also offered on Google Cloud’s Agent Platform (formerly Vertex AI), which has its own catalogue and lifecycle; Gemini 3.7 Flash, for example, is listed there as generally available. This page covers the Gemini Developer API, so check Google Cloud’s model pages before using these codes there.
Try a model for free in the Gemini API playground with a free Gemini API key, see which models your key can call with the API key checker, or compare prices in the model comparison.
Examples
Google says most production apps should use a specific stable model code such as gemini-3.8-flash. Preview codes can be shut down with two weeks’ notice, and -latest aliases are hot-swapped to each new release.
List your models: GET https://generativelanguage.googleapis.com/v1beta/models
| Model | API ID | Context |
|---|---|---|
| Gemini 3.8 Flash | 1,048,576 | |
| Gemini 3.5 Flash-Lite | 1,048,576 |
| Shuts down | Model IDs | Replacement |
|---|---|---|
| From 7 May 2027 | gemini-3.1-flash-lite | gemini-3.5-flash-lite |
FAQ
gemini-3.8-flash. It is a stable code, the kind Google recommends for production, with an input limit of 1,048,576 tokens and an output limit of 65,536 tokens. Google recommends it, or Gemini 3.5 Flash-Lite (gemini-3.5-flash-lite), for new projects.
It points to the latest release of the Flash model, whether stable, preview or experimental, and Google hot-swaps it with each new release. Google gives two weeks’ email notice before a breaking change. Its models page uses the alias only as an example, so call the models list endpoint to see what your key can use, and pin a stable code for production.
Google says preview models may be used for production, but they can have tighter rate limits and can be deprecated with as little as two weeks’ notice. If you use one, such as gemini-3.1-pro-preview, watch Google’s deprecations page and keep a stable fallback code ready.
Google limits the Gemini 2.5 models to users who have actively used them before, to keep capacity for older workloads. They are not deprecated, but new projects are pointed to Gemini 3.8 Flash or Gemini 3.5 Flash-Lite instead.
Google shut down gemini-2.0-flash and gemini-2.0-flash-001 on 1 Jun 2026, so requests to them fail. The recommended replacement is gemini-3.6-flash; for Gemini 2.0 Flash-Lite it is gemini-3.1-flash-lite.
Call GET https://generativelanguage.googleapis.com/v1beta/models with your key, or loop over client.models.list() in the Python SDK. Each result’s name is the model code to use. Google notes the name you pass should match one returned by models.list.
Related