JiaxuanLiu
JiaxuanLiu
About
Education
Experience
Projects
Articles
Awards
Contact
Light
Dark
Automatic
Speech Synthesis
Fun-CineForge: A Unified Dataset Pipeline and MLLM-based Model for Zero-Shot Movie Dubbing in Diverse Cinematic Scenes
A unified end-to-end pipeline (CineDub dataset) and MLLM-based model for zero-shot movie dubbing across diverse cinematic scenes — open-sourced via Alibaba Tongyi.
Jiaxuan Liu
PDF
Project
GitHub
Hugging Face
ModelScope
DiffStyleTTS: Diffusion-based Hierarchical Prosody Modeling for Text-to-Speech with Diverse and Controllable Styles
A diffusion-based acoustic model with hierarchical prosody modeling and improved classifier-free guidance, jointly enabling diverse and controllable prosody for TTS.
Jiaxuan Liu
,
Zhaoci Liu
,
Yajun Hu
,
Yingying Gao
,
Shilei Zhang
,
Zhenhua Ling
PDF
Project
Cite
×