Skip to content
AI Dev Toolkit.
Esc
  • AI Token CounterCount tokens for GPT, Claude, Gemini, DeepSeek, Qwen and more.Tool
  • LLM API Cost CalculatorEstimate per-request, daily and monthly API costs.Tool
  • AI Model ComparisonCompare prices, context windows and features across models.Tool
  • AI Model Pricing PagesSpecs, real costs and cheaper alternatives for popular models.Tool
  • Context Window CheckerSee whether your text fits each model’s context window.Tool
  • Subscription vs API CalculatorFind out whether a chat plan or the API is cheaper for you.Tool
  • GPU / VRAM CalculatorCheck how much VRAM a local model needs and which GPUs fit.Tool
  • Prompt Caching CalculatorEstimate savings from prompt caching.Tool

Model ID reference

Gemini model IDs for the Gemini API

Copy the exact Gemini API model code, from gemini-3.8-flash to preview models, with input and output limits, status and Google’s shutdown dates. Free, and it runs in your browser.

16 models. Click an ID to copy it.

Model IDs

  • Gemini 3.8 Flash

    CurrentRecommended

    1,048,576 context · 65,536 max output

    API ID

    • Stable
  • Gemini 3.5 Flash-Lite

    CurrentRecommended

    1,048,576 context · 65,536 max output

    API ID

    • Stable
  • Gemini 3.1 Pro Preview

    CurrentPreview

    1,048,576 context · 65,536 max output

    API ID

    • Preview
    • Preview
    Source
  • Gemini 3.6 Flash

    Legacy

    1,048,576 context · 65,536 max output

    API ID

    • Stable

    Replacement: gemini-3.8-flash

    Google calls it its previous-generation Flash model.

    Source
  • Gemini 3.7 Flash

    Legacy

    API ID

    • Redirect

    Requests to gemini-3.7-flash are automatically routed to gemini-3.8-flash.

  • Gemini 3.5 Flash

    Legacy

    API ID

    • Redirect

    Requests to gemini-3.5-flash are automatically routed to gemini-3.6-flash, although Google’s quickstart still uses this code in its examples.

  • Gemini 3 Flash Preview

    LegacyPreview

    1,048,576 context · 65,536 max output

    API ID

    • Preview

    Replacement: gemini-3.6-flash

    No shutdown date announced; Google recommends gemini-3.6-flash.

    Source
  • Gemini 2.5 Pro

    Legacy

    1,048,576 context · 65,536 max output

    API ID

    • Stable

    Not deprecated, but Google limits the 2.5 models to accounts that have used them before. New projects should use Gemini 3.8 Flash or 3.5 Flash-Lite.

    Source
  • Gemini 2.5 Flash

    Legacy

    1,048,576 context · 65,536 max output

    API ID

    • Stable

    Not deprecated, but Google limits the 2.5 models to accounts that have used them before. New projects should use Gemini 3.8 Flash or 3.5 Flash-Lite.

    Source
  • Gemini 2.5 Flash-Lite

    Legacy

    1,048,576 context · 65,536 max output

    API ID

    • Stable

    Not deprecated, but Google limits the 2.5 models to accounts that have used them before. New projects should use Gemini 3.8 Flash or 3.5 Flash-Lite.

    Source
  • Gemini 3.1 Flash-Lite

    Deprecated

    1,048,576 context · 65,536 max output

    API ID

    • Stable

    Shuts down 7 May 2027 at the earliest. Replacement: gemini-3.5-flash-lite

    Source
  • Gemini 2.0 Flash

    Retired

    API ID

    • Stable
    • Stable

    Retired 1 Jun 2026. Replacement: gemini-3.6-flash

  • Gemini 2.0 Flash-Lite

    Retired

    API ID

    • Stable
    • Stable

    Retired 1 Jun 2026. Replacement: gemini-3.1-flash-lite

  • Gemini 3.1 Flash-Lite Preview

    RetiredPreview

    API ID

    • Preview

    Retired 25 May 2026. Replacement: gemini-3.1-flash-lite

  • Gemini 2.5 Flash-Lite Preview

    RetiredPreview

    API ID

    • Preview

    Retired 31 Mar 2026. Replacement: gemini-3.1-flash-lite

  • Gemini 3 Pro Preview

    RetiredPreview

    API ID

    • Preview

    Retired 9 Mar 2026. Replacement: gemini-3.1-pro-preview

Steps

How to find and copy a Gemini model ID

  1. Search for a model (“flash”, “pro”, “2.5”) or paste a code from your app to see its status.
  2. Click a code to copy it. Use a Stable code in production, as Google recommends.
  3. Check codes marked Redirect: they still answer, but a different model serves them.
  4. Open “Show code” for the Gemini SDK call with that model code in place.

Method

How it works

Google calls the model ID a “model code”: the string you pass as model to the Gemini API, such as gemini-3.8-flash. This page lists the text models on the Gemini Developer API (Google AI Studio keys) from Google’s models and deprecations pages. Image, audio, music, video and embedding models are left out.

Stable, preview, latest and experimental codes

Google sorts every code into one of four version patterns:

  • Stable codes point to a specific model and usually don’t change. Google says most production apps should use one, such as gemini-3.8-flash.
  • Preview codes, such as gemini-3.1-pro-preview, may be used in production, typically have billing enabled, may have tighter rate limits and can be deprecated with at least two weeks’ notice.
  • Latest aliases, such as gemini-flash-latest, are hot-swapped with every new release of a model variation. Google emails two weeks’ notice before a breaking change behind one.
  • Experimental codes aren’t suitable for production and can disappear at any time.

Google notes this naming scheme dates from September 2025, so older models may follow different patterns: use the exact code from the list.

Which code to use

For new projects Google recommends Gemini 3.8 Flash (gemini-3.8-flash) or Gemini 3.5 Flash-Lite (gemini-3.5-flash-lite). Both take up to 1,048,576 input tokens and write up to 65,536 output tokens. The Gemini 2.5 models are not deprecated, but Google now limits them to accounts that have used them before, so a new key may not be able to call gemini-2.5-pro.

Codes that now route to another model

Two recent codes are kept alive by redirection: requests to gemini-3.7-flash go to gemini-3.8-flash, and requests to gemini-3.5-flash go to gemini-3.6-flash. Your code keeps working, but you get a different model from the one you named, with its own behaviour. Google’s quickstart still uses gemini-3.5-flash in its examples, so copy-pasted code may already be redirected.

Shutdown dates

Google’s deprecations page lists the earliest date each model might be shut down and says the exact date is confirmed with advance notice. Gemini 3.1 Flash-Lite, for example, can shut down from 7 May 2027 and the replacement is gemini-3.5-flash-lite. Once a model is shut down its endpoint stops responding: gemini-2.0-flash was switched off on 1 Jun 2026, with gemini-3.6-flash as the replacement.

Gemini on Google Cloud

Gemini models are also offered on Google Cloud’s Agent Platform (formerly Vertex AI), which has its own catalogue and lifecycle; Gemini 3.7 Flash, for example, is listed there as generally available. This page covers the Gemini Developer API, so check Google Cloud’s model pages before using these codes there.

Try a model for free in the Gemini API playground with a free Gemini API key, see which models your key can call with the API key checker, or compare prices in the model comparison.

Examples

Worked examples

Which Gemini ID to use in production

  • Google (Gemini)

    Google says most production apps should use a specific stable model code such as gemini-3.8-flash. Preview codes can be shut down with two weeks’ notice, and -latest aliases are hot-swapped to each new release.

    List your models: GET https://generativelanguage.googleapis.com/v1beta/models

Recommended Gemini models

The models Google currently recommends, with the ID to copy. Checked 11 Oct 2026.
ModelAPI IDContext
Gemini 3.8 Flash1,048,576
Gemini 3.5 Flash-Lite1,048,576

Upcoming Gemini shutdowns

Deprecated Gemini models by shutdown date, from Google’s deprecations page. Checked 11 Oct 2026.
Shuts downModel IDsReplacement
From 7 May 2027gemini-3.1-flash-litegemini-3.5-flash-lite

FAQ

Frequently asked questions

What is the model ID for Gemini 3.8 Flash?

gemini-3.8-flash. It is a stable code, the kind Google recommends for production, with an input limit of 1,048,576 tokens and an output limit of 65,536 tokens. Google recommends it, or Gemini 3.5 Flash-Lite (gemini-3.5-flash-lite), for new projects.

What does gemini-flash-latest point to?

It points to the latest release of the Flash model, whether stable, preview or experimental, and Google hot-swaps it with each new release. Google gives two weeks’ email notice before a breaking change. Its models page uses the alias only as an example, so call the models list endpoint to see what your key can use, and pin a stable code for production.

Can I use Gemini preview models in production?

Google says preview models may be used for production, but they can have tighter rate limits and can be deprecated with as little as two weeks’ notice. If you use one, such as gemini-3.1-pro-preview, watch Google’s deprecations page and keep a stable fallback code ready.

Why can’t my new API key use Gemini 2.5 Pro?

Google limits the Gemini 2.5 models to users who have actively used them before, to keep capacity for older workloads. They are not deprecated, but new projects are pointed to Gemini 3.8 Flash or Gemini 3.5 Flash-Lite instead.

What happened to gemini-2.0-flash?

Google shut down gemini-2.0-flash and gemini-2.0-flash-001 on 1 Jun 2026, so requests to them fail. The recommended replacement is gemini-3.6-flash; for Gemini 2.0 Flash-Lite it is gemini-3.1-flash-lite.

How do I list the Gemini models my key can use?

Call GET https://generativelanguage.googleapis.com/v1beta/models with your key, or loop over client.models.list() in the Python SDK. Each result’s name is the model code to use. Google notes the name you pass should match one returned by models.list.