VFF - The signal in the noise
NewsTrending

Google Launches Gemini Omni for AI-Powered Video Generation and Editing

Read original
Share
Google Launches Gemini Omni for AI-Powered Video Generation and Editing

Google DeepMind has introduced Gemini Omni, a multimodal model that generates and edits video from mixed inputs including images, audio, video, and text. The first model in the family, Gemini Omni Flash, is rolling out to the Gemini app, Google Flow, and YouTube Shorts with the ability to edit videos through natural language conversation while maintaining character consistency and physical coherence across multiple turns. Future versions will support additional output modalities like image and audio generation.

  • Gemini Omni Flash enables video generation and editing from mixed input modalities (text, image, audio, video)
  • Users can edit videos conversationally with natural language, with edits building on previous instructions while maintaining scene consistency
  • Initial rollout targets Gemini app, Google Flow, and YouTube Shorts, with image and audio output modalities planned for future releases
  • The model grounds video generation in Gemini's real-world knowledge and allows users to transform existing footage or create entirely new content

Gemini Omni represents a significant step in multimodal AI capability, moving beyond text-to-image generation into video creation and editing. This consolidates reasoning and creative generation into a single model, which could reshape how creators and enterprises approach video production and editing workflows. The conversational editing interface lowers the technical barrier for complex video manipulation tasks.

For content creators and media companies, this tool could reduce production timelines and costs by enabling rapid iteration on video content through natural language prompts rather than traditional editing software. For Google, this positions Gemini as a competitive alternative to specialized video generation tools and integrates generative capabilities deeper into YouTube and its productivity suite.

  • Video generation and editing may shift from specialized software to conversational AI interfaces, affecting the competitive landscape for traditional video editing tools
  • Multimodal input handling at scale suggests progress toward more general-purpose AI systems that can reason across and generate across multiple content types
  • Integration into YouTube Shorts and Google Flow signals Google's strategy to embed generative capabilities into existing user-facing products rather than launching standalone tools

Monitor adoption rates and user feedback on video quality, consistency, and editing accuracy across multiple turns. Watch for competitive responses from other AI labs and video software vendors, and track whether Google expands output modalities (image, audio) on the timeline promised. Pay attention to any content moderation or authenticity challenges that emerge as video generation becomes more accessible.

Related Video

Share

Subscribe to the newsletter

The latest stories and analysis, delivered to your inbox.

Free. No spam. Unsubscribe any time.

Related stories

Relay shuts down, team joins Google Chrome
TrendingNews

Relay shuts down, team joins Google Chrome

AI automation startup Relay has shut down, with its staff joining Google's Chrome team. Founder and CEO Jacob Bank indicated the team will work on integrating AI capabilities into Chrome to help users accomplish tasks. The move represents Google's continued expansion of AI features across its product ecosystem.

by Lucas Ropek· TechCrunch AI
Google Cuts Gemini Flash Pricing 50% With Faster Iteration
TrendingModel Release

Google Cuts Gemini Flash Pricing 50% With Faster Iteration

Google DeepMind released Gemini 3.7 Flash on August 13, 2026, positioning it as an improved workhorse model for coding and agent-based tasks. The model arrives three weeks after Gemini 3.6 Flash and delivers measurable gains in software engineering, web development, and knowledge-intensive workflows at half the per-token cost of its predecessor. The release reflects developer feedback and algorithmic improvements aimed at production-ready code generation and complex document processing.

· Google Deepmind
Google DeepMind Launches Sign Language AI for Deaf Users
TrendingNews

Google DeepMind Launches Sign Language AI for Deaf Users

Google DeepMind has introduced sign-language-to-text (SL2T), a new AI model that converts sign language into text for Deaf and hard of hearing users. The model powers new sign language features designed to improve accessibility. The announcement marks a significant step in making AI tools more inclusive for sign language users.

· Google Deepmind
Google's Gemini Hits 1B Monthly Users
TrendingNews

Google's Gemini Hits 1B Monthly Users

Google's Gemini app has reached 1 billion monthly active users, according to CEO Sundar Pichai, making it the fastest-growing product in the company's history. The milestone signals Google's growing competitive position in the consumer chatbot market, which has been dominated by OpenAI's ChatGPT. The announcement was made Tuesday on X.

by Alix Coutures· The Information