Back to News

ByteDance Releases SeedRealtime: Full-Duplex Audio-Video Model for Natural Interaction

#bytedance#seedrealtime#full-duplex#multimodal

ByteDance's Seed team unveiled SeedRealtime, a native full-duplex audio-video large model that jointly understands sound, images, and temporal information. It enables real-time, natural interaction where the model can see, listen, and speak simultaneously, marking a step toward fully multimodal conversational AI.

Coverage timeline

  1. ByteDance Seed Blog

    作为原生音视频全双工大模型,联合理解声音、画面与时序信息,带来边看、边听、边说的自然交互体验