Post-Cutoff.com
  1. Home
  2. Timeline
  3. 2026
  4. ByteDance deploys SeedRealtime, a native audio-visual…

ByteDance deploys SeedRealtime, a native audio-visual full-duplex model, in the Doubao app

★★★after cutoffmodel-releaseByteDanceconfidence: high

ByteDance Seed launched SeedRealtime, an end-to-end LLM that listens, watches (live video) and speaks at the same time instead of chaining ASR, vision and TTS, and rolled it out at scale in Doubao (Dola internationally). ByteDance says it halves conversational pacing problems compared with cascaded systems.

Key facts

What happened

ByteDance's Seed team shipped an audio-visual full-duplex model to Doubao, China's largest consumer chatbot, letting users hold natural video-call-style conversations with the assistant (demos include menu translation, museum guiding and walking through an espresso machine).

Why it matters

It puts end-to-end "see, hear and talk at once" interaction in front of a mass consumer audience. ByteDance published only human-evaluation claims, no quantitative benchmarks.

Changelog

  • 2026-09-29: created

Models

Related events

  1. StepFun releases StepAudio 3 family; its Realtime model tops Artificial Analysis full-duplex rankings ★★★
  2. Alibaba launches Qwen-Audio-3.1 five-model voice stack and cuts audio API prices up to 95% ★★★
  3. ByteDance Seed LiveInterpret 2.0: end-to-end Chinese-English simultaneous interpretation in your own voice, ~3 s behind ★★★

Sources (4)

id: 2026-08-05-bytedance-seedrealtime · updated 2026-09-29 · open in the interactive timeline