PANews
PANews|Aug 05, 2026 04:53
**[ByteDance Releases SeedRealtime, a Full-Duplex Audio-Video Model]** According to a report by CLS, ByteDance has officially launched its native full-duplex audio-video model, SeedRealtime. SeedRealtime integrates audio, video, and text through a unified architecture, enabling real-time interaction across continuous multimodal information streams, delivering a new "watch, listen, and speak simultaneously" experience. End-to-end human evaluation results show that compared to cascade models, SeedRealtime reduces audio-video dialogue rhythm issues by half—demonstrating a more natural grasp of speaking timing. Issues such as "interrupting before the speaker finishes, delayed responses after the speaker finishes, or being mistakenly triggered by background noise or casual chatter" have been significantly reduced. Additionally, the probability of completing a single conversation smoothly and coherently has also seen a notable improvement. Currently, SeedRealtime has been fully rolled out on the Doubao App.
+1
Mentioned
Share To

Timeline

HotFlash

APP

X

Telegram

Facebook

Reddit

CopyLink

Hot Reads