aiminute. ← All AI news
Tools AI Minute Newsroom 2026-09-04

Meta's new model tells 20 speakers apart while they talk at once.

Meta's new model tells 20 speakers apart while they talk at once.

Meta released Muse Voice Transcribe on 1 September, its first real-time audio model. One pass writes the words, names the speaker and marks when someone stops, 0.16 seconds after speech ends. Meta sells it only as an API at three dollars per thousand audio minutes, with no downloadable weights.

Why it mattersLive captions and meeting notes have needed three stitched models to do this job. Folding it into one makes always-on transcription cheap enough to leave running.
#Audio & Music

✓ Verified · 4 sources

WhatsApp X Telegram
Read in the app — free, in 9 languages

Related stories

Runway's new model generates a playable world at 24 frames a second.
2026-09-05
Attackers are stealing OpenAI keys from a popular AI building tool.
2026-09-04
American doctors can now pull patient records into ChatGPT.
2026-09-03
Reliance Jio now streams a PC to any old computer in India.
2026-09-03
This tool hides AI writing by rewriting the plot, not the word choice.
2026-09-02