VFF - The signal in the noise
NewsTrending

ElevenLabs adds selective editing to music generation model

Read original
Share
ElevenLabs adds selective editing to music generation model

ElevenLabs has released a new music-generation model that allows users to regenerate specific sections of a song while preserving the rest of the track. The capability enables mid-track genre switching and targeted edits without requiring full song regeneration. This addresses a key limitation in current music AI tools that typically require regenerating entire compositions for any modifications.

  • ElevenLabs' new model supports selective regeneration of song sections
  • Users can switch genres mid-track without affecting other parts
  • Eliminates need to regenerate entire songs for targeted edits
  • Represents incremental improvement in music generation control and flexibility

Music generation AI has struggled with granular editing capabilities. This update gives creators more control over their outputs, reducing iteration friction and enabling more experimental workflows. The ability to modify sections independently makes AI-generated music more practical for professional and hobbyist use.

ElevenLabs competes in an increasingly crowded music AI space. Selective editing is a feature that could differentiate its offering and improve user retention by reducing the need to restart generation workflows. This positions the tool as more production-ready for musicians and content creators.

  • Music AI tools are moving toward finer-grained control, mimicking traditional DAW workflows
  • Selective regeneration could accelerate adoption among professional musicians skeptical of full-generation approaches
  • Competitive pressure will likely push other music AI platforms to match or exceed this capability

Monitor whether other music AI platforms (Suno, Udio, others) adopt similar selective editing features. Track adoption metrics from ElevenLabs to see if this capability drives meaningful user growth. Watch for integration with DAWs and music production software, which would signal broader industry acceptance.

Related Video

Share

Subscribe to the newsletter

The latest stories and analysis, delivered to your inbox.

Free. No spam. Unsubscribe any time.

Related stories

AWS Shows How to Build AI Phone Ordering for Restaurants

AWS Shows How to Build AI Phone Ordering for Restaurants

AWS published a technical walkthrough for building a voice-based restaurant ordering system using Amazon Connect, Lex V2, and AI agents. The system answers inbound calls, takes orders through natural conversation, and confirms them without requiring apps, websites, or customer logins. The architecture separates the telephony channel, conversation logic, and backend services to keep ordering functionality independent from how customers access it.

by Sergio Barraza· AWS Machine Learning Blog
How Top Speech Models Game Benchmarks
TrendingNews

How Top Speech Models Game Benchmarks

Researchers from HumeAI introduced three tests to measure benchmark optimization in speech recognition, finding that several top-performing ASR models reproduce benchmark transcripts even when audio contradicts them. Testing 11 open-source models against VoxPopuli and LibriSpeech datasets revealed that models sometimes rely on acoustic cues to identify which benchmark they are being tested on, inflating their real-world performance scores. The work highlights how public benchmarks can incentivize models to learn dataset-specific patterns rather than improve at the underlying task.

· Hugging Face Blog
Wispr raises $280M at $2B valuation, expands beyond dictation
TrendingNews

Wispr raises $280M at $2B valuation, expands beyond dictation

Wispr raised $280 million in new funding at a $2 billion valuation, bringing its total funding to over $361 million. The funding round signals investor confidence in the voice AI company as it expands beyond dictation use cases. The company is positioning itself for growth in a competitive market for speech recognition and voice-based AI applications.

by Ivan Mehta· TechCrunch AI
LTX-2.5 Generates Video Faster Than Real-Time, Pushes Open Weights Forward

LTX-2.5 Generates Video Faster Than Real-Time, Pushes Open Weights Forward

LTX released LTX-2.5, an open-weights video generation model that produces 10-second clips in 6.8 seconds on Nvidia GB200 chips, with native multishot support and improved quality. The model is available free for organizations under $10 million ARR on Hugging Face, ComfyUI, and via API. LTX claims 33 million downloads across its model family and reports a 67% win rate in blind quality tests against competing models.

by carl.franzen@venturebeat.com (Carl Franzen)· VentureBeat AI