aiminute. ← All AI news
New Models AI Minute Newsroom 2026-08-05

ByteDance's SeedRealtime watches, listens and talks — all at the same time

ByteDance's SeedRealtime watches, listens and talks — all at the same time

ByteDance's Seed team released SeedRealtime on Wednesday, a model that natively fuses audio, video and text in one end-to-end system for full-duplex interaction: it keeps watching and listening even while it speaks, instead of taking turns. The company says it reads sound, visuals and timing together to work out who is addressing it and what they want, and it is already live in the Doubao assistant app. It extends April's voice-only Seeduplex model to the visual world.

Why it mattersFull-duplex voice was step one; adding eyes makes an AI assistant behave less like a walkie-talkie and more like someone in the room. ByteDance shipping this straight into its flagship consumer assistant signals that real-time multimodal interaction is becoming the next battleground for AI apps.
#Video#Audio & Music

✓ Verified · 2 sources

WhatsApp X Telegram
Read in the app — free, in 9 languages

Related stories

Reflection will hand out a 501-billion-parameter model for free.
2026-10-06
Microsoft's new transcriber starts writing before you finish speaking.
2026-10-05
ElevenLabs is giving students a free year of its reading voice.
2026-10-05
An American lab is about to give away a China-class open model.
2026-10-05
A video site now gives away the best-scoring open translation model.
2026-10-05