Gemini 3.5 Flash-Lite

Google

Our fastest, most cost-effective 3.5-class model, delivering 350 output tokens per second.

Specs

Model

Family
Gemini
Variant
3.5 Flash-Lite
Weights
Proprietary

Limits

Context
1,048,576 tokens
Max output
65,536 tokens

Modalities

Input
Text Image Video Audio PDF
Output
Text

Pricing

Input
$0.3 / 1M
Output
$2.5 / 1M
Cache read
$0.03 / 1M

Capabilities

Reasoning
Yes
Reasoning control
Effort levels
Tool calling
Yes
Attachments
Yes

Dates

Released
21 Jul 2026
Knowledge cutoff
Mar 2026
Google

Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber

We’re introducing new Gemini models, including Gemini 3.6 Flash, 3.5 Flash-Lite and 3.5 Flash Cyber.

Tulsee Doshi
blog.google

More from Google

All models

Gemini 3.5 Flash-Cyber

Google

Fine-tuned for finding and fixing cybersecurity vulnerabilities at a lower price per token than larger models.

Gemini 3.6 Flash

Google

Our workhorse model that delivers better coding, knowledge work, and multimodal performance.

$1.5 / $7.5 per 1M · $0.15 cached 23d ago

Keep up with the tools

An occasional email when notable AI dev tools and models land in the directory. No spam, unsubscribe anytime.