Model spec · 2026-08-30

Gemini 3.5 Flash-Lite
Standard $0.30 / $2.50

As of 2026-08-30 the official pricing table lists Standard $0.30 input / $2.50 output per million tokens (output includes thinking). Stable id gemini-3.5-flash-lite, 1,048,576 input, 65,536 output. This page does not claim the id is live on QCode.

#gemini-3.5-flash-lite#$0.30 / $2.50#1,048,576#high-throughput

Takeaways

$0.30 / $2.50

Standard price / 1M tokens

Official Standard: $0.30 input (text/image/video/audio), $2.50 output including thinking. Versus 3.5 Flash at $1.50 / $9.00, input is about 5× cheaper.

1,048,576

Input limit

Official model table Input token limit, same as 3.5 Flash.

65,536

Output limit

Official model table Output token limit. Some overview pages round to 64,000; this page uses 65,536 from the model table.

2026-07

Model page latest update

Official model page Latest update: July 2026. Versions: Stable gemini-3.5-flash-lite.

What 3.5 Flash-Lite is

Google describes a low-latency, cost-efficient multimodal Flash for high-throughput sub-agents, document parsing, and simple extraction. Inputs: text, image, video, audio, PDF. Output: text. The capability table is close to 3.5 Flash: caching, code execution, function calling, Search grounding, thinking. Computer use is Preview. No audio generation, image generation, or Live API.

What people still search

Lite searches are billing questions, not launch news. 3.5 Flash Standard is $1.50 / $9.00; Lite is $0.30 / $2.50. Do not mix the id with gemini-3.1-flash-lite — that is a different, cheaper Flash-Lite line ($0.25 text/image/video input). This page is only 3.5 Flash-Lite.

Timeline

2026-07

Model page Latest update: July 2026. Stable id gemini-3.5-flash-lite.

Pricing table

Still listed on 2026-08-30: Standard $0.30 / $2.50. Batch / Flex $0.15 / $1.25. Priority $0.54 / $4.50.

2026-08-30

This fetch of the model page and pricing table. 3.6 / 3.7 intro rates belong to a different Flash SKU — do not paste them onto Lite.

Confirmed vs watch-outs

Confirmed

Stable id gemini-3.5-flash-lite; 1,048,576 / 65,536; Standard $0.30 / $2.50. Sources: official model page and pricing, fetched 2026-08-30.

Watch-outs

Do not write 3.5 Flash-Lite as 3.1 Flash-Lite. Window numbers follow the model table 65,536, not overview pages that say 64,000. This page does not say it is already available on QCode or that you can switch on the same endpoint.

Lite vs 3.5 Flash

3.5 Flash-Lite (this page)

High throughput, lower price, sub-agents or extraction. Price it at $0.30 / $2.50.

3.5 Flash

When you need more reasoning depth, use gemini-3.5-flash at Standard $1.50 / $9.00. Input is about 5× the Lite rate.

How to call it

On Gemini API / AI Studio / Vertex set model=gemini-3.5-flash-lite. Follow the model page for capabilities. Treat Computer use as Preview. Do not casually swap in the 3.1 id.

On QCode

This page does not claim gemini-3.5-flash-lite is already callable on QCode, and it does not say you can switch on the same endpoint. Whether the id appears in the live catalog is /models plus your own probe. For Google’s own path, use the Gemini API.

FAQ

What is the model id?

gemini-3.5-flash-lite. The official model page marks it Stable.

What does Google charge now?

Standard: $0.30 input / $2.50 output per million tokens, output including thinking. Batch / Flex $0.15 / $1.25. Priority $0.54 / $4.50. Fetched 2026-08-30.

How much cheaper than 3.5 Flash?

3.5 Flash Standard is $1.50 / $9.00. Lite input is about 1/5; output is $2.50 vs $9.00.

Is the window 64k or 65,536?

The model table says Output token limit 65,536. Some overview pages say 64,000. This page follows the model table.

Can it generate images or audio?

No. Inputs may be image, video, audio, and PDF; output is text. Image generation, audio generation, and Live API are unsupported.

Can I call it on QCode?

This page makes no availability promise. Use the live /models list and your own test. Do not read a catalog snapshot as “already available.”

Sources

ai.google.dev model page /gemini-api/docs/models/gemini-3.5-flash-lite (Last updated 2026-07-30 UTC); pricing /gemini-api/docs/pricing Gemini 3.5 Flash-Lite section. 3.5 Flash prices from the same pricing page. Fetched 2026-08-30.

Read Lite’s official price first

3.5 Flash-Lite specs follow Google’s docs. QCode’s catalog is the live /models list.

Related

Not affiliated with Google / DeepMind. Prices and windows follow ai.google.dev. This page is not a promise that QCode can call this id.