Google (Gemini) Gemini 3.7 Flash — what it is, when to use it and what to use instead

Active

No. Gemini 3.7 Flash is active at Google (Gemini) and has no announced end date.

gemini-3.7-flash last verified

Summary

Gemini 3.7 Flash is the previous release in Google's Flash line, at $0.75 per million input tokens and $3.75 per million output tokens. It shares its limits, modalities and price with the newer Gemini 3.8 Flash.

Gemini 3.7 Flash is listed with an August 2026 update and was followed in September 2026 by Gemini 3.8 Flash. Because the two carry the same $0.75 / $3.75 price, the same 1,048,576-token input limit and the same 65,536-token output limit, the newer model is the default choice unless you have a reason to pin this one.

It takes text, image, video, audio and PDF input and returns text, with structured outputs, caching, function calling, code execution, search grounding, URL context, the Batch API and thinking at low, medium or high. The Live API is not supported, and it does not generate images or audio.

No deprecation has been announced.

Modality
  • Text
  • Vision
Use case
  • Chat
  • Coding
  • Reasoning
  • Agents & tool use
  • Long context
Tier
  • Mid-tier
Access
  • Hosted API

Specifications

Input price
$0.75 / M tokens
Output price
$3.75 / M tokens
Price band
Standard
Context window
1M tokens
Max output
65.5K tokens
Input
Text, image
Output
Text
Knowledge cutoff
unknown
Tool use
Yes
Structured output
Yes
Weights licence
not published
Model maker
Google (Gemini)

“unknown” means the figure has not been published, or we have not verified it yet — never that the answer is no. “not published” under the weights licence means this model is not offered with downloadable weights. The price band is ours: it bands the blended price (three parts input to one part output) as budget (≤ $0.50), standard (≤ $3), premium (≤ $12) or flagship.

Lifecycle

Google (Gemini) has not announced a deprecation or shutdown date for this model. That is a statement about what has been published, not a guarantee — providers typically announce a shutdown months ahead, and this page updates when they do.

  1. Released
  2. Deprecated not announced
  3. Retired not announced

Use it for

  • Keep an existing Gemini 3.7 Flash integration running while you evaluate Gemini 3.8 Flash against it.
  • Pin a specific Flash generation where output has to stay comparable across releases.
  • Run multimodal agentic workloads that combine text with images, video, audio or PDFs.

Avoid it when

  • You are starting a new integration; Gemini 3.8 Flash costs the same and is a generation newer.
  • You need more than 65,536 output tokens in a single response.
  • You need a live, low-latency session; the Live API is not supported on this model.

Alternatives to consider

  • Gemini 3.8 Flash Curated

    Google (Gemini) Active

    The newer Flash release at an identical $0.75 / $3.75 with the same limits and modalities.

    $0.75 in · $3.75 out / M tokens 1M context

  • Gemini 3.6 Flash Curated

    Google (Gemini) Active

    The earlier Flash release at the same $0.75 / $3.75, if you need to compare against the generation before this one.

    $0.75 in · $3.75 out / M tokens 1M context

  • Claude Sonnet 5 Curated

    Anthropic Active

    Anthropic's mid tier at $2 / $10 with a 1M-token window and 128K max output.

    $2.00 in · $10.00 out / M tokens 1M context

  • GPT-5.6 Terra Curated

    OpenAI Active

    OpenAI's balanced tier at $2 / $12 with a 1.05M-token window and a February 2026 knowledge cutoff.

    $2.00 in · $12.00 out / M tokens 1.1M context

  • Gemini 3.7 Flash Derived

    DeepInfra Active

    Alternative from DeepInfra, same tier (mid-tier), similar price; shares chat, coding, reasoning, agents and long-context.

    $0.75 in · $3.75 out / M tokens 1M context

Curated suggestions were picked by a reviewer, with the reason written by hand. Derived suggestions are ranked from the data — shared tags, the same tier and a comparable price band — not from benchmark scores. Check the capability fit yourself: we track lifecycle, price and specifications.

Sources

Verified against the source on . Dataset snapshot . Providers can change a date without notice — the link above is always the final word.

About this profile

The specifications and dates above are facts from Google (Gemini), with the sources listed. The summary, the use-for and avoid-when lists and the curated alternatives are our assessment: editorial profile reviewed on by llm-researcher.

Common questions

What is Gemini 3.7 Flash for?
Gemini 3.7 Flash is the previous release in Google's Flash line, at $0.75 per million input tokens and $3.75 per million output tokens. It shares its limits, modalities and price with the newer Gemini 3.8 Flash.
What does Gemini 3.7 Flash cost?
$0.75 per million input tokens and $3.75 per million output tokens, as last verified on 2026-09-10.
What is the context window of Gemini 3.7 Flash?
1M tokens.
Is Gemini 3.7 Flash deprecated?
No. Gemini 3.7 Flash is active at Google (Gemini) and has no announced end date.
What should I use instead of Gemini 3.7 Flash?
Google (Gemini) Gemini 3.8 Flash — The newer Flash release at an identical $0.75 / $3.75 with the same limits and modalities. Google (Gemini) Gemini 3.6 Flash — The earlier Flash release at the same $0.75 / $3.75, if you need to compare against the generation before this one. Anthropic Claude Sonnet 5 — Anthropic's mid tier at $2 / $10 with a 1M-token window and 128K max output.

All Google (Gemini) models and their end dates →