Gemini 3.5 Flash-Lite

Google

Our fastest, most cost-effective 3.5-class model, delivering 350 output tokens per second.

Specs

Model

Family
Gemini
Variant
3.5 Flash-Lite
Weights
Proprietary

Limits

Context
1,048,576 tokens
Max output
65,536 tokens

Modalities

Input
Text Image Video Audio PDF
Output
Text

Pricing

Input
$0.3 / 1M
Output
$2.5 / 1M
Cache read
$0.03 / 1M

Capabilities

Reasoning
Yes
Reasoning control
Effort levels
Tool calling
Yes
Attachments
Yes

Dates

Released
21 Jul 2026
Knowledge cutoff
Mar 2026

Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber

We’re introducing new Gemini models, including Gemini 3.6 Flash, 3.5 Flash-Lite and 3.5 Flash Cyber.

Tulsee Doshi
blog.google

More from Google

All models

Gemini 3.8 Flash Cyber

Google

Our most capable cybersecurity model with frontier-level performance in vulnerability detection and automated patching, available to trusted defenders through our new Fairwind Program.

—25d ago

Gemini 3.8 Flash

Google

Our most intelligent workhorse model, delivering significant improvements from 3.7 Flash across software engineering, agentic tasks, and critical, multi-step reasoning in specialized domains.

$0.75 / $3.75 per 1M · $0.075 cached 25d ago

Gemini Omni 1.1 Flash

Google

Omni now delivers studio-quality video production, including the ability to extend a scene, first and last frame interpolation, crisp 4K upscaling, faster prototyping, and more.

—4w ago

Keep up with the tools

An occasional email when notable AI dev tools and models land in the directory. No spam, unsubscribe anytime.