Nari Labs Dia2 (1B / 2B)
English only. Successor to Dia-1.6B (April 2025, github.com/nari-labs/dia). Release date 2025-11-19 from secondary sources (GitHub releases page).
- Input
- text, audio
- Output
- audio
- License
- apache-2.0
- Verified
- 2026-09-29
How to call it
| Provider | Model id | Endpoint / URL | Docs |
|---|---|---|---|
| Hugging Face | nari-labs/Dia2-2B | huggingface.co/nari-labs/Dia2-2B | — |
| GitHub | — | github.com/nari-labs/dia2 | — |
Notable capabilities (1)
- Streaming multi-speaker dialogue TTS: Generates [S1]/[S2] dialogue and starts producing audio from the first few input tokens (no need for full text); conditions on audio prefixes for real-time conversation; up to ~2 min per generation (Mimi codec, 12.5 Hz); word-level timestamps. source