VFF - The signal in the noise
Model ReleaseTrending

Google Launches Gemini 3.5 Transcribe with 85+ Language Support

Read original
Share
Google Launches Gemini 3.5 Transcribe with 85+ Language Support

Google has released Gemini 3.5 Transcribe, a new transcription model that automatically detects specialized jargon and supports more than 85 languages. The model represents an improvement over its predecessor, Chirp 3, with better multilingual performance and lower wording error rates. Users can edit transcriptions using voice commands. The release comes as Google continues to roll out updates to its Gemini Audio suite while the promised Gemini 3.5 Pro model remains unreleased since its June launch window.

  • Google launched Gemini 3.5 Transcribe with support for 85+ languages and automatic jargon detection
  • Model shows improvements in multilingual performance and wording error rates versus Chirp 3
  • Voice-based editing capabilities allow users to modify transcriptions naturally
  • Release follows 3.5 Live Translate launch, but Gemini 3.5 Pro remains unavailable since June

Transcription accuracy and multilingual support are critical for global teams and knowledge workers who rely on audio-to-text conversion for meetings, interviews, and content creation. Better jargon detection reduces post-transcription cleanup work, while voice-based editing removes friction from the correction workflow. These improvements address real pain points in how professionals capture and process spoken information.

Organizations using Google's transcription tools can reduce manual editing time and improve accuracy across multilingual teams without switching platforms. The voice-editing feature lowers the barrier to correcting transcripts, potentially increasing adoption and reducing reliance on manual transcription services. However, the delayed Gemini 3.5 Pro release may signal resource constraints or product prioritization shifts at Google.

  • Transcription workflows become faster and less labor-intensive for teams that adopt voice-based editing
  • Global organizations can rely on a single tool for 85+ languages rather than managing multiple transcription services
  • Automatic jargon detection may reduce errors in specialized fields like medicine, law, and engineering
  • Delayed Gemini 3.5 Pro release raises questions about Google's product roadmap and competitive positioning

Monitor whether Gemini 3.5 Transcribe adoption accelerates among enterprise customers and whether the voice-editing feature reduces post-transcription editing time in practice. Track the eventual release of Gemini 3.5 Pro and whether it includes transcription capabilities or represents a different product direction. Watch for competitive responses from other transcription providers like Otter.ai, Rev, and OpenAI's Whisper.

Related Video

OneUpAI
OneUp Your Business. Get More Done. OneUp Your Business. Get More Done. OneUp Your Business. Get More Done.
Learn More
Share

Subscribe to the newsletter

The latest stories and analysis, delivered to your inbox.

Free. No spam. Unsubscribe any time.

Related stories

Particle's Radar makes 130K podcasts searchable for AI

Particle's Radar makes 130K podcasts searchable for AI

Particle has launched a podcast intelligence platform called Radar that transcribes and analyzes over 130,000 podcasts, making their content searchable on the web and accessible to AI agents via API and MCP. The platform enables both human users and AI systems to query podcast conversations at scale. This addresses a significant gap in AI training data and search accessibility for audio content.

by Sarah Perez· TechCrunch AI
MiniMax's ARR Surges to $800M on Enterprise AI Demand
TrendingNews

MiniMax's ARR Surges to $800M on Enterprise AI Demand

Chinese AI developer MiniMax reported its annualized revenue run rate surged to $800 million in August, up from $150 million in February, driven by enterprise demand for its AI video model and large language models. The Shanghai-based company's growth reflects a broader shift toward commercializing AI capabilities for business customers. The expansion demonstrates significant market traction in China's competitive AI sector.

by Juro Osawa· The Information
Ringg raises $10M to expand voice AI beyond phone calls
TrendingNews

Ringg raises $10M to expand voice AI beyond phone calls

Ringg, an India-based voice AI company, has raised $10 million from Peak XV Partners as part of a Series A extension. The funding supports the company's efforts to expand voice AI applications beyond traditional phone calls. The investment reflects growing investor interest in voice-based AI solutions in emerging markets.

by Ivan Mehta· TechCrunch AI
AWS Shows How to Build AI Phone Ordering for Restaurants

AWS Shows How to Build AI Phone Ordering for Restaurants

AWS published a technical walkthrough for building a voice-based restaurant ordering system using Amazon Connect, Lex V2, and AI agents. The system answers inbound calls, takes orders through natural conversation, and confirms them without requiring apps, websites, or customer logins. The architecture separates the telephony channel, conversation logic, and backend services to keep ordering functionality independent from how customers access it.

by Sergio Barraza· AWS Machine Learning Blog