DLLM-TTS: Block discrete diffusion language model for text-to-speech synthesis
Read the original at arxiv.org→arXiv:2608.00011v1 Announce Type: new Abstract: Current text-to-speech systems face a trade-off: autoregres- sive codec language models produce highly intelligible speech but require large-scale models and training...
Original headline: "DLLM-TTS: Block Discrete Diffusion Language Model for Text-to-Speech Synthesis"
Coverage timeline
- Aug 4, 04:00 UTC arXiv cs.CL lead source DLLM-TTS: Block Discrete Diffusion Language Model for Text-to-Speech Synthesis