Skip to content

Gemini 2.0 Flash

Gemini 2.0 Flash was Google's fast, cost-efficient large language model in the Gemini 2 generation, designed for high-throughput, latency-sensitive applications. Google has since released several newer Flash models in the Gemini 3.x generation, and Gemini 2.0 Flash no longer appears on Google's current Gemini API pricing page.

By Google DeepMindPay as you goSince 2015AI Models

Facts checked on against the sources listed at the end of this page.

Newer version available. Google has released newer Flash models (Gemini 3.5, 3.6, 3.7 and 3.8 Flash) since Gemini 2.0 Flash, which no longer appears on Google's current Gemini API pricing page.

Pricing
Pay as you goPaid from Superseded by current Gemini API pricing
Open source
NoProprietary

A good fit for

  • Researching the history and progression of Google's Gemini Flash model line
  • Understanding capability and pricing context for Google's newer Gemini 3.x Flash models

Not a good fit for

  • New projects, since Google has released several newer Flash models since Gemini 2.0 Flash
  • Anyone wanting Google's latest available fast, cost-efficient Gemini model

Gemini 2.0 Flash pricing

Gemini 2.0 Flash was priced per million tokens through the Gemini API. It no longer appears on Google's current Gemini API pricing page, which now features newer Gemini 3.x Flash models; check Google's current pricing page for active models.

Prices in USD, checked on 27 September 2026. Current prices on ai.google.dev.

Main features

Low-latency responses
Built for high-throughput, latency-sensitive applications.
API access
Available through the Gemini API, Google AI Studio and Vertex AI.
Cost-efficient tier
Positioned below Gemini's flagship Pro-tier models on price.

About Gemini 2.0 Flash

Gemini 2.0 Flash was positioned as Google's fast, cost-efficient model within the Gemini 2 generation, aimed at high-throughput and latency-sensitive applications where a lighter, cheaper model than Gemini's flagship Pro tier was preferred.

It was accessible through the Gemini API, Google AI Studio and Vertex AI, and was used broadly in production applications valuing speed and cost efficiency over maximum capability.

Google has since released several newer Flash-tier models in the Gemini 3.x generation, including 3.5, 3.6, 3.7 and 3.8 Flash, and Gemini 2.0 Flash no longer appears on Google's current Gemini API pricing page.

Made by
Google DeepMind
Free plan
Access was available through a free tier with rate limits, historically
Available on
API, Google AI Studio, Vertex AI, Gemini apps (historically)

Gemini 2.0 Flash alternatives

Questions about Gemini 2.0 Flash

Is Gemini 2.0 Flash still available?

It no longer appears on Google's current Gemini API pricing page, which now features newer Gemini 3.x Flash models; availability should be confirmed directly with Google.

What replaced Gemini 2.0 Flash?

Google has released several newer Flash-tier models since then, including Gemini 3.5, 3.6, 3.7 and 3.8 Flash.

Sources and links

Work on Gemini 2.0 Flash? Claim this listing. Something wrong or out of date? Suggest a correction. Stats come from the public GitHub and npm APIs and update every day.

Gemini 2.0 Flash: pricing, features and alternatives (2026) | ScratchDB