
Xue Liumeng is a tenure-track Assistant Professor at the School of Intelligence Science and Technology, Nanjing University, and an Executive Member of the Speech, Dialogue, and Auditory Processing Technical Committee of the China Computer Federation (CCF). Her research focuses on audio, speech, and language processing, including speech, music, and audio understanding and generation, as well as emotional and expressive speech understanding and generation. She received her Ph.D. from Northwestern Polytechnical University, during which she conducted research at JD AI Lab, Tencent AI Lab, and Microsoft, leading to industrial applications and two patent filings. She subsequently conducted postdoctoral research at The Chinese University of Hong Kong, Shenzhen and The Hong Kong University of Science and Technology.She is a co-founder of Amphion, a comprehensive audio generation and visualization platform integrating text-to-speech, singing voice conversion, text-to-sound generation, and visual analytics, which has repeatedly trended on GitHub and received broad attention from international communities and media. She led the development of Audio-FLAN, a unified instruction-following dataset for speech, audio and music understanding and generation, which ranked second on the Hugging Face Dataset Trending list and has been downloaded over 100,000 times by top technology companies and research groups, including Google, Meta, Nvidia, ByteDance, Tencent, Alibaba, Cambridge, Oxford, Tsinghua, ETH Zurich AI Center, and LAION. She has contributed to large audio language model research, including Llasa, Spark-TTS, YuE, and AudioX, actively promoting open-source applications of generative audio technologies. Her work has been published in top conferences and journals such as ACL, ICLR, ICASSP, INTERSPEECH, IEEE/ACM TASLP, and Neural Networks, and has attracted attention and citations from leading technology companies, including Apple and Amazon.She has served on the organizing committees of international conferences including ISCSLP 2026 and IEEE SLT 2024, as well as workshops and challenges such as MLC-SLM@INTERSPEECH 2026, SmartGlasses@SLT 2026, LLM4MA Workshop@ISMIR 2025, and CoVoC@ISCSLP 2024. She is a regular reviewer for top-tier conferences and journals, including ACL, ACM MM, ICASSP, INTERSPEECH, IEEE/ACM TASLP, Speech Processing Letters, and Speech Communication. She maintains close collaborations with leading industry companies and research institutions and actively supports students in internships at top technology companies and in international laboratory exchanges.
Personal homepage:https://lmxue.github.io/

Xue Liumeng is a tenure-track Assistant Professor at the School of Intelligence Science and Technology, Nanjing University, and an Executive Member of the Speech, Dialogue, and Auditory Processing Technical Committee of the China Computer Federation (CCF). Her research focuses on audio, speech, and language processing, including speech, music, and audio understanding and generation, as well as emotional and expressive speech understanding and generation. She received her Ph.D. from Northwestern Polytechnical University, during which she conducted research at JD AI Lab, Tencent AI Lab, and Microsoft, leading to industrial applications and two patent filings. She subsequently conducted postdoctoral research at The Chinese University of Hong Kong, Shenzhen and The Hong Kong University of Science and Technology.She is a co-founder of Amphion, a comprehensive audio generation and visualization platform integrating text-to-speech, singing voice conversion, text-to-sound generation, and visual analytics, which has repeatedly trended on GitHub and received broad attention from international communities and media. She led the development of Audio-FLAN, a unified instruction-following dataset for speech, audio and music understanding and generation, which ranked second on the Hugging Face Dataset Trending list and has been downloaded over 100,000 times by top technology companies and research groups, including Google, Meta, Nvidia, ByteDance, Tencent, Alibaba, Cambridge, Oxford, Tsinghua, ETH Zurich AI Center, and LAION. She has contributed to large audio language model research, including Llasa, Spark-TTS, YuE, and AudioX, actively promoting open-source applications of generative audio technologies. Her work has been published in top conferences and journals such as ACL, ICLR, ICASSP, INTERSPEECH, IEEE/ACM TASLP, and Neural Networks, and has attracted attention and citations from leading technology companies, including Apple and Amazon.She has served on the organizing committees of international conferences including ISCSLP 2026 and IEEE SLT 2024, as well as workshops and challenges such as MLC-SLM@INTERSPEECH 2026, SmartGlasses@SLT 2026, LLM4MA Workshop@ISMIR 2025, and CoVoC@ISCSLP 2024. She is a regular reviewer for top-tier conferences and journals, including ACL, ACM MM, ICASSP, INTERSPEECH, IEEE/ACM TASLP, Speech Processing Letters, and Speech Communication. She maintains close collaborations with leading industry companies and research institutions and actively supports students in internships at top technology companies and in international laboratory exchanges.
Personal homepage:https://lmxue.github.io/