JiaxuanLiu
JiaxuanLiu
About
Education
Experience
Projects
Articles
Awards
Contact
Light
Dark
Automatic
Classifier-Free Guidance
DiffStyleTTS: Diffusion-based Hierarchical Prosody Modeling for Text-to-Speech with Diverse and Controllable Styles
A diffusion-based acoustic model with hierarchical prosody modeling and improved classifier-free guidance, jointly enabling diverse and controllable prosody for TTS.
Jiaxuan Liu
,
Zhaoci Liu
,
Yajun Hu
,
Yingying Gao
,
Shilei Zhang
,
Zhenhua Ling
PDF
Project
Cite
×