Preprint
Aug 2026
VoiceChat-TTS: A Low-Latency Continuous Speech Synthesis Model for Interactive Agents
VoiceChat-TTS is proposed, a low-latency, continuous, and streamable text-to-speech model for interactive agents that enables always-on, responsive speech generation while preserving modularity and high speech quality.
Edresson Casanova, Jaehyeon Kim, Mariana Graterol Fuenmayor et al.
· 3 citations