Gemini 3.5 Flash-Lite
Standard $0.30 / $2.50
As of 2026-08-30 the official pricing table lists Standard $0.30 input / $2.50 output per million tokens (output includes thinking). Stable id gemini-3.5-flash-lite, 1,048,576 input, 65,536 output. This page does not claim the id is live on QCode.
Takeaways
Standard price / 1M tokens
Official Standard: $0.30 input (text/image/video/audio), $2.50 output including thinking. Versus 3.5 Flash at $1.50 / $9.00, input is about 5× cheaper.
Input limit
Official model table Input token limit, same as 3.5 Flash.
Output limit
Official model table Output token limit. Some overview pages round to 64,000; this page uses 65,536 from the model table.
Model page latest update
Official model page Latest update: July 2026. Versions: Stable gemini-3.5-flash-lite.
What 3.5 Flash-Lite is
Google describes a low-latency, cost-efficient multimodal Flash for high-throughput sub-agents, document parsing, and simple extraction. Inputs: text, image, video, audio, PDF. Output: text. The capability table is close to 3.5 Flash: caching, code execution, function calling, Search grounding, thinking. Computer use is Preview. No audio generation, image generation, or Live API.
What people still search
Lite searches are billing questions, not launch news. 3.5 Flash Standard is $1.50 / $9.00; Lite is $0.30 / $2.50. Do not mix the id with gemini-3.1-flash-lite — that is a different, cheaper Flash-Lite line ($0.25 text/image/video input). This page is only 3.5 Flash-Lite.
Timeline
Model page Latest update: July 2026. Stable id gemini-3.5-flash-lite.
Still listed on 2026-08-30: Standard $0.30 / $2.50. Batch / Flex $0.15 / $1.25. Priority $0.54 / $4.50.
This fetch of the model page and pricing table. 3.6 / 3.7 intro rates belong to a different Flash SKU — do not paste them onto Lite.
Confirmed vs watch-outs
Confirmed
Stable id gemini-3.5-flash-lite; 1,048,576 / 65,536; Standard $0.30 / $2.50. Sources: official model page and pricing, fetched 2026-08-30.
Watch-outs
Do not write 3.5 Flash-Lite as 3.1 Flash-Lite. Window numbers follow the model table 65,536, not overview pages that say 64,000. This page does not say it is already available on QCode or that you can switch on the same endpoint.
Lite vs 3.5 Flash
3.5 Flash-Lite (this page)
High throughput, lower price, sub-agents or extraction. Price it at $0.30 / $2.50.
3.5 Flash
When you need more reasoning depth, use gemini-3.5-flash at Standard $1.50 / $9.00. Input is about 5× the Lite rate.
How to call it
On Gemini API / AI Studio / Vertex set model=gemini-3.5-flash-lite. Follow the model page for capabilities. Treat Computer use as Preview. Do not casually swap in the 3.1 id.
On QCode
This page does not claim gemini-3.5-flash-lite is already callable on QCode, and it does not say you can switch on the same endpoint. Whether the id appears in the live catalog is /models plus your own probe. For Google’s own path, use the Gemini API.
FAQ
What is the model id?
gemini-3.5-flash-lite. The official model page marks it Stable.
What does Google charge now?
Standard: $0.30 input / $2.50 output per million tokens, output including thinking. Batch / Flex $0.15 / $1.25. Priority $0.54 / $4.50. Fetched 2026-08-30.
How much cheaper than 3.5 Flash?
3.5 Flash Standard is $1.50 / $9.00. Lite input is about 1/5; output is $2.50 vs $9.00.
Is the window 64k or 65,536?
The model table says Output token limit 65,536. Some overview pages say 64,000. This page follows the model table.
Can it generate images or audio?
No. Inputs may be image, video, audio, and PDF; output is text. Image generation, audio generation, and Live API are unsupported.
Can I call it on QCode?
This page makes no availability promise. Use the live /models list and your own test. Do not read a catalog snapshot as “already available.”
Sources
ai.google.dev model page /gemini-api/docs/models/gemini-3.5-flash-lite (Last updated 2026-07-30 UTC); pricing /gemini-api/docs/pricing Gemini 3.5 Flash-Lite section. 3.5 Flash prices from the same pricing page. Fetched 2026-08-30.
Read Lite’s official price first
3.5 Flash-Lite specs follow Google’s docs. QCode’s catalog is the live /models list.
Related
Gemini 3.5 Flash guide
Non-Lite Flash in the same generation, $1.50 / $9.00.
Gemini 3.6 Flash guide
A newer Flash SKU.
Gemini 3.7 Flash pricing
How to read intro pricing.
Not affiliated with Google / DeepMind. Prices and windows follow ai.google.dev. This page is not a promise that QCode can call this id.