Topic
Voice & Video AI
Text-to-speech, voice cloning, video generation, and audio AI
Featured
All Stories
Google Launches Gemini 3.5 Transcribe with 85+ Language Support
Google has released Gemini 3.5 Transcribe, a new transcription model that automatically detects specialized jargon and…
Particle's Radar makes 130K podcasts searchable for AI
Particle has launched a podcast intelligence platform called Radar that transcribes and analyzes over 130,000 podcasts,…

AWS Shows How to Build AI Phone Ordering for Restaurants
AWS published a technical walkthrough for building a voice-based restaurant ordering system using Amazon Connect, Lex…

How Top Speech Models Game Benchmarks
Researchers from HumeAI introduced three tests to measure benchmark optimization in speech recognition, finding that…
Wispr raises $280M at $2B valuation, expands beyond dictation
Wispr raised $280 million in new funding at a $2 billion valuation, bringing its total funding to over $361 million.…

LTX-2.5 Generates Video Faster Than Real-Time, Pushes Open Weights Forward
LTX released LTX-2.5, an open-weights video generation model that produces 10-second clips in 6.8 seconds on Nvidia…
Ford launches AI assistant for vehicle info in mobile app
Ford is launching an AI-powered chatbot assistant in its Ford and Lincoln mobile apps that can answer questions about…
Smallest.ai raises $13M for human-sounding voice AI
Smallest.ai has raised $13 million in funding to develop voice AI models designed to conduct phone calls that pass the…

How Runway Turned a Bug Into a Feature
Runway ML shared insights on building real-time AI video models at VB Transform 2026, revealing that the company turned…
Fish Audio Raises $52M on Strong Voice AI Traction
Fish Audio, an AI voice model startup, raised $52 million in seed funding after launching last year with 8 million…
OpenAI brings voice mode to ChatGPT desktop app
OpenAI has rolled out voice mode to its ChatGPT desktop application, enabling users to interact with ChatGPT Work and…

Black Forest Labs Launches FLUX 3 Video Model in Limited Release
Black Forest Labs launched FLUX 3, a multimodal AI model capable of generating images and video with audio up to 20…
Runway launches model router to simplify generative media selection
Runway has launched Media Router, an automated tool that selects the optimal image, video, or audio generation model…
Anthropic Brings Voice Mode to Stronger Claude Models
Anthropic has expanded voice mode access beyond its Haiku model to include Opus and Sonnet, its more capable AI models.…
Google Vids adds AI avatars for personalized video creation
Google has added personalized AI avatars to its Vids product, enabling users to create videos featuring digital…

Cars24 scales to 1M monthly conversations with OpenAI agents
Cars24, an automotive marketplace, deployed OpenAI-powered voice and chat agents to automate customer conversations at…
Hinge Founder Raises $18M for AI Voice Dating App Overtone
Justin McLeod, founder of dating app Hinge, has raised $18 million to launch Overtone, a new AI-powered dating service…
Spotify Adds ChatGPT-Like Assistant for Premium Discovery
Spotify is launching a ChatGPT-like conversational AI assistant for Premium subscribers that allows users to discover…

ScienceSoft builds HIPAA-compliant AI voice scheduler on AWS
ScienceSoft has built a HIPAA-compliant AI voice scheduler using Amazon Nova Sonic and Amazon Bedrock Guardrails on AWS…

Deutsche Telekom Deploys OpenAI Across Operations
Deutsche Telekom is integrating OpenAI technology to transform its operations across customer service, employee…
Character.AI enters microdrama with AI-animated video series
Character.AI announced c.ai Series, a new product combining short-form episodic videos with interactive elements…

SpaceXAI Pushes Into Consumer AI With Grok 4.5 and Voice Agents
SpaceXAI and its soon-to-be subsidiary Cursor jointly introduced Grok 4.5, a coding and agentic task model that Elon…

