VFF - The signal in the noise
Model Release

Google Launches Custom Voice Generation with Gemini 3.8 TTS

Read original
Share
Google Launches Custom Voice Generation with Gemini 3.8 TTS

Google introduced Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, new text-to-speech models that enable users to create custom voices from scratch using natural language prompts. The Flash model supports granular control over performance details like pacing, emotion, and dialect, while the Flash-Lite variant prioritizes cost-efficient, high-volume applications. Both models are available across Google AI Studio, Gemini API, Gemini Enterprise, Gemini Notebook, and Google Vids, with built-in watermarking for security.

  • Google launched two new text-to-speech models: Gemini 3.8 Flash TTS for creative character design and Gemini 3.8 Flash-Lite TTS for high-volume, cost-efficient use cases
  • Users can create custom voices from scratch across more than 100 languages and dialects using natural language prompts, scaling from 30 original voices to an infinite library
  • The Flash model offers line-by-line directional control over acting cues, pacing, dialect shifts, and backchanneling for immersive audiobooks, games, and podcasts
  • Models include built-in watermarking and are integrated into Google's broader Gemini Audio family alongside Live Translate, Transcribe, and Live Extended Thinking

Text-to-speech technology has historically relied on static, pre-built voice presets. These models shift the paradigm toward dynamic, on-demand voice generation with fine-grained creative control, lowering barriers for content creators and enterprises to produce audio at scale without hiring voice actors or managing talent.

For enterprises and developers, these models reduce production costs and timelines for audiobooks, podcasts, dubbing, and voice agents while enabling personalized brand voices. The cost-optimized Flash-Lite variant specifically targets high-volume use cases where expressive quality matters but per-unit cost is critical.

  • Voice acting and dubbing workflows may shift from human talent to AI-generated custom voices, affecting labor demand in audio production
  • Content creators can now produce multilingual, expressive audio content without geographic or linguistic constraints, enabling faster global content distribution
  • Watermarking and built-in safety tools suggest Google is positioning these models as enterprise-grade, but detection and misuse risks remain open questions

Monitor adoption rates across Google Vids, Gemini Notebook, and third-party integrations via the Gemini API to gauge real-world demand. Watch for regulatory responses to synthetic voice generation, particularly around consent, deepfakes, and voice cloning in jurisdictions with emerging AI governance frameworks.

OneUpAI
OneUp Your Business. Get More Done. OneUp Your Business. Get More Done. OneUp Your Business. Get More Done.
Learn More
Share

Subscribe to the newsletter

The latest stories and analysis, delivered to your inbox.

Free. No spam. Unsubscribe any time.

Related stories

Google DeepMind Adds Private Memory to AI Compute
TrendingNews

Google DeepMind Adds Private Memory to AI Compute

Google DeepMind has introduced private, server-side memory capabilities for its Private AI Compute offering, designed to enable personal AI applications while maintaining data privacy. The advancement allows AI models to access and utilize memory on secure servers without exposing user data to the broader system. This development addresses a key technical challenge in deploying private AI systems that require persistent context while maintaining cryptographic isolation.

· Google Deepmind
YouTube Lets Users Build Custom Feeds With Gemini AI
TrendingNews

YouTube Lets Users Build Custom Feeds With Gemini AI

YouTube is introducing custom feeds that allow users to describe the videos they want to see in natural language, with Gemini AI building personalized feeds based on those descriptions. The feature gives users direct control over feed curation rather than relying solely on YouTube's algorithmic recommendations. This represents a shift toward user-directed content discovery on the platform.

by Sarah Perez· TechCrunch AI
Google DeepMind Opens AGI Institute to Broaden Debate
TrendingNews

Google DeepMind Opens AGI Institute to Broaden Debate

Google DeepMind has launched a new institute designed to surface and debate differing perspectives on artificial general intelligence (AGI) between Google, Google DeepMind, and the global research community. The institute acknowledges that stakeholders will not always agree and may change positions as new data emerges in the rapidly evolving AGI field. The move signals an effort to broaden the conversation around AGI development beyond internal company views.

by Aditya Mehta· TechCrunch AI
Google Opens Smart Home to Third-Party AI Agents
TrendingNews

Google Opens Smart Home to Third-Party AI Agents

Google is opening Google Home to third-party AI agents through a new integration called Home MCP, which uses the standardized Model Context Protocol. The move allows agents like Claude, Open Claw, Google Antigravity, and Hermes to securely access, control, and monitor connected devices and analyze home data within the Google Home ecosystem. This represents a shift toward interoperability in smart home control, letting users choose which AI agent manages their connected devices.

by Jennifer Pattison Tuohy· The Verge AI