Gemini 3.8 text-to-speech says hello
Gemini 3.8 Flash-Lite TTS and Gemini 3.8 Flash TTS are our most expressive audio models yet.
Built for deep creative direction and character design. Create entirely new voices from scratch using natural language prompts to bring characters to life across gaming, immersive audiobooks, podcasts, and interactive media. Direct every performance line by line with granular control over acting cues, pacing, dialect shifts, and backchanneling.
Built for high-volume, cost-efficient scale. Optimized for high-volume dubbing, audio content creation, and expressive voice agents with fine-grained control over tone, pacing, and expressive nuance.
Built for high-complexity tasks, with increased intelligence and multi-step reasoning.
Built for scale and cost efficiency, combining conversational intelligence with fluid dialogue and visual grounding.
An occasional email when notable AI dev tools and models land in the directory. No spam, unsubscribe anytime.