In the news
Gemini 3.8 text-to-speech says hello
Google DeepMind · Published · 3 min read
In 30 seconds
- What happened
- Google released Gemini 3.8 Flash TTS and Flash-Lite TTS models that generate custom voices from natural language prompts with line-by-line performance control.
- Why it matters
- Developers building voice agents, audiobooks, podcasts, or dubbed content need expressive, scalable text-to-speech with fine-grained emotional and pacing control.
- Watch out
- Voice replication requires consent verification, but long-term effectiveness of SynthID watermarking against misuse remains unproven in real-world deployment.
- speech
- text-to-speech
- gemini
The patterns behind this
Each one covers how the technique works, when it earns its cost, and where it breaks.
The Agent Architect
One pattern, one tradeoff, one production failure story. A short weekly briefing for people building agentic systems.
Weekly email, one-click unsubscribe. We only use your address to send the briefing.