See how ChatGPT, Perplexity and Google AI Overviews describe you today

Free AI Visibility Audit
Google model°

Gemini Flash Latest

Gemini Flash Latest is a Google alias that always points at the newest model in the Gemini Flash family, so an application calling it moves to each new Flash release without a code change.

GoogleVerified 24/09/2026Released 27/04/2026

01 / decision

Decision snapshot

Each figure sits next to the middle of the field, so it can be read as dear or cheap, wide or narrow, rather than floating on its own. Compared against the 231 models we track, not against an absolute standard.

Capability

Not measured

No published benchmark scores yet

Input / 1M tokens

$0.75

field median $0.43

Output / 1M tokens

$3.75

field median $1.80

Context

1,049K

field median 500K

02 / overview

What Gemini Flash Latest is, and when to reach for it

Four questions, answered separately, because somebody arrives at one of them rather than at the top of the page.

What it is

The job it was built for, and the job it is not for.

Gemini Flash Latest is not a fixed model, it is a rolling pointer to whichever Gemini Flash model Google currently ships as newest. The job it does is keep an integration current, so a team building high-volume multimodal work on Flash inherits the latest version automatically. It is the wrong choice when you need a stable, pinned target whose behaviour will not shift under you, for example anything under evaluation, audit or regression testing.

When it arrived, and when to use it

Where it sits in its line, and when a sibling is the better pick.

Gemini Flash Latest was first seen in April 2026. It sits alongside the pinned Gemini Flash releases as the always-current entry point to that family, where Flash itself is Google's fast, cheaper tier rather than its heaviest reasoning tier. Pick it when you want new Flash capability the day it lands and can absorb a change in behaviour, pick a pinned Flash version when you cannot.

How you reach it

The API, the apps it powers, and what its limits let you do.

You reach Gemini Flash Latest through Google's Gemini API by naming the alias instead of a specific version, and requests then resolve to the current Flash model. It accepts text, images, video, audio and file uploads, so it will take a support inbox, a long recording, a screen capture or a PDF in the same call. The context window runs to seven figures of tokens and output can run long, which means whole document sets, transcripts or codebases go in at once rather than being chunked and stitched.

Why it matters

What changes because this exists, or why it does not.

The practical change is maintenance. Teams running Flash at volume have usually pinned a version and then scheduled the migration work later, and this alias removes that chore for anyone whose workload tolerates version drift. It is otherwise unremarkable in capability terms, it is the Flash family behind a stable name rather than a new set of abilities.

03 / evidence

How much of this is verified

Split by category, so a strong number never hides a thin evidence base. Verified means we read it on the benchmark's own published results; a provider's claim about its own model is shown and labelled rather than dropped.

No published benchmark scores for this model yet.

We publish a score only where we can link the result to where it was published. Until a benchmark result for this model exists in a source we read, this section stays empty rather than being filled with a provider's marketing figure.

How we decide what counts as evidence

04 / ledger

Benchmark ledger

Every published row, grouped by category, each compared with the best published score on the same benchmark. 'Is 64% good' is a question nobody can answer; '26 points behind the leader' is one anybody can.

Nothing in the ledger yet.

Each row here carries a score, the benchmark it came from, what the leading model scored on the same test, and a link to the published result. Rows appear as results are published and read.

How we decide what counts as evidence

05 / capability

Capability shape

Where this model is strong, and against how many peers. Ranks are against models with evidence in that category, not against everything we track: ranking against models nobody tested would rank who published, not who is better.

No category scores to shape yet.

A category score is the weighted mean of the benchmarks published for it. With no published rows there is nothing to average, and an empty chart drawn at zero would say something false.

How we decide what counts as evidence

06 / cost

What it costs

List API rates as last read from the provider, with the source on every row, plus every change we have recorded since we started tracking it.

Gemini Flash Latest API pricingSurge45°
ChargePriceUnitRead onSource
Input$0.75per 1M tokens2026-09-24Check
Output$3.75per 1M tokens2026-09-24Check

Running it costs the same as the Flash tier it resolves to, which sits at the low end of frontier pricing and is built for jobs you run millions of times rather than once. Output is charged at several times the input rate, so the economics favour long inputs with short, structured answers, and a costlier pinned flagship only earns its place on the work Flash visibly struggles with.

What a month costsSurge45°
WorkloadInput / monthOutput / monthCost
A small product team20M tokens5M tokens$33.75
A busy support assistant200M tokens40M tokens$300.00
A document pipeline1000M tokens100M tokens$1,125.00

List API rates, no caching and no batch discount, which both providers offer and which change the answer a great deal. Treat these as the ceiling, not the bill.

07 / specs

Specifications

As listed by the provider's own catalogue and re-read every few hours. Anything absent is absent there too.

SpecificationSurge45°
Context window1,048,576 tokens
Maximum output65,536 tokens
Modalitiestext, image, video, file, audio
Released27/04/2026
StatusCurrent
Catalogue identifier~google/gemini-flash-latest

08 / lineage

Lineage

What this model replaced, what replaced it, and what else its provider has in the field.

09 / line

The line

Every model in this family in release order, so a page from eight months ago says in one glance that two newer ones exist.

Nothing else in this line yet.

A line is worked out from the naming and the release dates across every model page we hold. It fills in as the provider ships successors, or as we pick up the models that came before this one.

How we decide what counts as evidence

10 / notes

Our notes

What this model changes for a brand trying to be cited in AI answers, and every change we have logged since it launched.

What it changes for you

Because Flash is the cheap, fast tier, it handles a large share of the high-volume AI answers your buyers actually see, and an alias like this means that tier updates underneath you without notice. For a SaaS brand, that argues against tuning content to one model's quirks and for the durable things every Flash version rewards: clear product pages, named use cases, specifics a model can lift and attribute. Treat citation checks as continuous rather than a one-off audit, since the model answering your category question this month may not be the one answering it next.

Where buyers meet this model

Buyers rarely meet the alias by name. They meet the Flash model underneath it in the Gemini consumer app, in AI Overviews and AI Mode in Google Search, and in Workspace features, while developers meet the alias directly in the Gemini API when they want their app to track the newest Flash release.

Change log

Nothing published here yet. Changes appear within hours of a provider announcing them.

11 / questions

Questions

The things people ask about this model, answered from what is on this page rather than from anywhere else.

Is Gemini Flash Latest a separate model?
No. It is an alias that redirects to the newest model in the Gemini Flash family, so what answers your call is whichever Flash release is current at that moment.
When should I pin a Flash version instead?
Whenever behaviour needs to stay fixed: evaluations, regression suites, regulated workflows, or any prompt chain you have tuned carefully. The alias can change under you, a pinned version will not.
What can I send it?
Text, images, video, audio and file uploads, all through the same API call, with a context window large enough to take whole document sets or long recordings without chunking.
Surge45°

Is Gemini Flash Latest recommending you?

Models change what gets cited. We measure whether AI answers name your brand or your competitors across every assistant, and show you what to change.