Model board / Gemini 3.5 Flash-Lite

Gemini 3.5 Flash-Lite

latest Google · Gemini

Verified 2026-09-10 against the pages listed below. Prices are USD per 1M tokens, list rate — no committed-use, enterprise or volume discount.

Input
$0.3
Cached input
$0.03
Cache write
Output
$2.50
Context window
Max output
Weights
closed
Caching
both

one input rate for text, image, video and audio — no audio premium, where Gemini 3.1 Flash-Lite charges $0.50 for audio against $0.25 for text. Cache storage $1.00 per 1M/hr

not stated on the pricing page or the models list

implicit caching is on by default for Gemini 2.5 and newer; explicit caches also bill storage per hour. The caching page's minimum-prefix table does not list this model

Watch out

Dearer than Gemini 3.1 Flash-Lite on text in ($0.30 against $0.25) and on everything out ($2.50 against $1.50), and cheaper only on audio in. Google calls it the most cost-effective 3.5 model, which is a claim about the 3.5 line, not about the board

What we would use it for

Audio-heavy input, where it bills audio at the text rate

What this has cost, week by week

Output price at each refresh since 2026-09-10. A vendor change is the provider changing its price. Our correction means this board changed what it prints and the provider's charge did not move — the two are never merged, because publishing the second as the first would be a price rise that never happened.

09-10 $2.50

Nothing has moved since we started tracking it.

Where these numbers came from

No number on this page was written that was not read from one of those pages on 2026-09-10. If one is wrong, send us the source URL and it is corrected in the next weekly refresh, with a changelog line saying what was wrong — contact@naderu.com.

Jobs this is one of our picks for

Discussed in Naderu Weekly

The rest of the board

This row sits alongside 57 others on the model board, refreshed every week. The same data as JSON, and the terms for using it, are at /board-data.