ByteDance Releases SeedRealtime: Full-Duplex Audio-Video Model for Natural Interaction
#bytedance#seedrealtime#full-duplex#multimodal
ByteDance's Seed team unveiled SeedRealtime, a native full-duplex audio-video large model that jointly understands sound, images, and temporal information. It enables real-time, natural interaction where the model can see, listen, and speak simultaneously, marking a step toward fully multimodal conversational AI.
Coverage timeline
ByteDance Seed Blog
作为原生音视频全双工大模型,联合理解声音、画面与时序信息,带来边看、边听、边说的自然交互体验
