aiminute. ← All AI news
New Models AI Minute Newsroom 2026-09-14

Xiaomi's open model picks one voice out of three talking at once.

Xiaomi's open model picks one voice out of three talking at once.

Xiaomi released the free CocktailASR-1 speech model on Friday. You give it a short sample of the voice you want, and it ignores everyone else. On a three-voice test it got 12.29% of words wrong, against 4.11% with two voices.

Why it mattersMeeting notes and captions fall apart the moment two people talk over each other. Open weights mean this can run on your own machine, not a company's server.
#Audio & Music

✓ Verified · 3 sources

▶ Related video: Solving the Cocktail Party Problem with Deep Learning (In-depth version)
WhatsApp X Telegram
Read in the app — free, in 9 languages

Related stories

DeepSeek is testing its first voice mode on a slice of users.
2026-09-14
ElevenLabs' new music model lets free users sell what it writes.
2026-09-13
A Chinese retailer's free model draws the world as you walk into it.
2026-09-13
Kimi's cheap coding model now opens a million-token window to all.
2026-09-12
A Chinese map app built a 3D world you can fly through.
2026-09-11