I am a first-year Master’s student at the Northwestern Polytechnical University, supervised by Prof. Lei Xie.

My research interest includes speech synthesis and song generation.

🔍 Research Area

  • Speech Processing: Text-to-Speech
  • Music Generation: Song Generation, Singing Voice Synthesis, Music Information Retrieval

🎓 Education

M.S. in Computer Science
Northwestern Polytechnical University (NWPU)
Audio, Speech and Language Processing Group (ASLP@NPU) 2025.9 – Present
B.S. in Computer Science
Northwestern Polytechnical University (NWPU)
School of Computer Science 2021.9 – 2025.7

📝 Publications

First / Co-First Author

CCF-B Interspeech 2026
YingMusic-Singer: Controllable Singing Voice Synthesis with Flexible Lyric Manipulation and Annotation-free Melody Guidance
C Hao, J Zheng, G Ma, Y Jiang, H Chen, W Tian, G Chen, Z Chen, L Xie.
CCF-B ICME 2026
SongFormer: Scaling Music Structure Analysis with Heterogeneous Supervision
C Hao, R Yuan, J Yao, Q Deng, X Bai, W Xue, L Xie.

Co-Author

CCF-B ICME 2026 Grand Challenge (Efficiency Track 1st Place)
S2Accompanist: A Semantic-Aware and Structure-Guided Diffusion Model for Music Accompaniment Generation
H Chen, W Cheng, G Ma, C Hao, Y Xia, M Wei, Z Zhao, P Zhu, H Zhang, L Xie.
CCF-C ASRU 2025
DiffRhythm+: Controllable and Flexible Full-Length Song Generation with Preference Optimization
H Chen, Y Jiang, G Ma, C Hao, S Wang, J Yao, Z Ning, M Meng, J Luan, L Xie.
Preprint
DiffRhythm: Blazingly Fast and Embarrassingly Simple End-to-End Full-Length Song Generation with Latent Diffusion
Z Ning, H Chen, Y Jiang, C Hao, G Ma, S Wang, J Yao, L Xie.
Preprint
SongEval: A Benchmark Dataset for Song Aesthetics Evaluation
J Yao, G Ma, H Xue, H Chen, C Hao, Y Jiang, H Liu, R Yuan, J Xu, W Xue, H Liu, L Xie.

🛠️ Open Source Projects

Open Source
ArxivWatcher
Automated daily arXiv paper monitoring with LLM-powered analysis, web browsing, RSS feeds, Zotero import, and optional email/Feishu notifications.