Biography

Zhe Li is a Postdoctoral Fellow at The University of Hong Kong. His research interests include speech LLMs, robust speaker representation learning, and multimodal artificial intelligence for healthcare applications. He received his Ph.D. in Electrical and Electronic Engineering from The Hong Kong Polytechnic University in 2025, his M.Sc. in Software Engineering from Xinjiang University in 2021, and his B.Eng. in Computer Science from Qilu University of Technology in 2016. He was a research intern at Microsoft Research Asia (MSRA) in 2025 and a visiting Ph.D. researcher in the Department of Electrical Engineering at Stanford University in 2024. He has led two research projects and contributed to a project funded by the Hong Kong Research Grants Council. He has published more than 60 papers in leading speech journals and conferences, including IEEE TASLP, ICASSP, and INTERSPEECH. He holds three granted invention patents and one software copyright. He delivered tutorials on speech large language models at ICME 2026 and INTERSPEECH 2026, and co-organized a special session at ACM MMAsia 2026. He received the 2020 Outstanding Scientific and Technological Achievement Award from the Chinese Association for Artificial Intelligence. His co-authored work received the Best Student Paper Runner-Up Award at PRICAI 2024.

“You are more than what you have become!”

News

More
  • Jun. 2026: 🎉 4 papers have been accepted to INTERSPEECH 2026! Many thanks to all collaborators and co-authors for their great work — see you September 27–October 1 in Sydney, Australia! 🇦🇺
  • Apr. 2026: 🎉 Our paper “DB-SMGA: Dual-Branch Sequential Multi-Granularity Attention for Speech Depression Detection” has been accepted for publication in IEEE Signal Processing Letters (SPL). Congratulations to Dr. Meirong Song for her excellent work!
  • Apr. 2026: 🎉 Our paper “Uncertainty-Aware Multi-Head Multi-Mode Knowledge Distillation for Self-Supervised Speaker Verification” has been accepted by IEEE Transactions on Audio, Speech, and Language Processing (TASLP)! Thanks to Dr. Jin!
  • Apr. 2026: 🎉 Our tutorial Speech Large Language Models for Under-Resourced Languages has been accepted by INTERSPEECH 2026 — see you September 27–October 1 in Sydney, Australia! 🇦🇺
  • Mar. 2026: 🎉 Our paper Towards A Unified Perspective on Parameter-Efficient Fine Tuning for Speaker Verification has been accepted by IEEE Transactions on Audio, Speech, and Language Processing (TASLP)! Thanks to Prof. Mak!
  • Jan. 2026: 🎉 Two papers accepted to ICASSP 2026 — see you May 4–8 in Barcelona, Spain! 🇪🇸
  • Dec. 2025: 🎉 My First Tutorial! Our tutorial Speech Large Language Models: Architectures, Efficient Adaptation, and Applications has been accepted by IEEE ICME 2026 — see you July 5–9 in Bangkok, Thailand! 🇹🇭
  • Sep. 29, 2025: 🎉 Our paper “WhisMultiNet: Advancing End-to-End Speech Topic Classification with Whisper and MultiGateGNN” has been accepted by IEEE Transactions on Audio, Speech, and Language Processing (TASLP)! Thanks to Xiaozhe Qi!
  • Sep. 4, 2025: 🎉 Our paper “Disentangling Speech Representations Learning with Latent Diffusion for Speaker Verification” has been accepted by IEEE Transactions on Audio, Speech, and Language Processing (TASLP)! Thanks to Prof. Mak!
  • Aug. 20, 2025: 🎉 One paper accepted to EMNLP 2025 — see you in Suzhou, China! 🇨🇳
  • Jun. 18, 2025: 🎉 One paper accepted to MICCAI 2025 — see you in Daejeon, South Korea! 🇰🇷
  • Jun. 14, 2025: 🎉 Our paper “Mutual Information-Enhanced Contrastive Learning with Margin for Maximal Speaker Separability” has been accepted by IEEE Transactions on Audio, Speech, and Language Processing (TASLP). Thanks to Prof. Mak!
  • May 19, 2025: 🎉 Two papers accepted to INTERSPEECH 2025 — see you in Rotterdam, the Netherlands! 🇳🇱
  • Mar. 4, 2025: 🧑🏻‍🏫 Paper Sharing Session: I gave a talk on Spectral-Aware Low-Rank Adaptation for Speaker Verification (ICASSP 2025).
  • Feb. 11, 2025: 🧑🏻‍💻 Joined Microsoft Research Asia (MSRA) as a Research Intern, focusing on multimodal large models for healthcare.
  • Dec. 21, 2024: 🎉 Four papers accepted to ICASSP 2025 — see you in Hyderabad, India! 🇮🇳
  • Dec. 4, 2024: 🏅 Enhancing Multimodal Rumor Detection with Statistical Image Features and Modal Alignment via Contrastive Learning received the Best Student Paper Runner-Up Award 🥈 at PRICAI 2024.
  • Jun. 17, 2024: 🧑🏻‍🏫 Paper Sharing Session: Parameter-efficient Fine-tuning of Speaker-Aware Dynamic Prompts for Speaker Verification (INTERSPEECH 2024).
  • Apr. 3, 2024: 🧑🏻‍🏫 Paper Sharing Session: Dual Parameter-Efficient Fine-Tuning for Speaker Representation via Speaker Prompt Tuning and Adapters (ICASSP 2024).
  • Dec. 8, 2023: Presented Maximal Speaker Separability via Robust Speaker Representation Learning at NCMMSC 2023, Soochow, China. 🇨🇳
  • Dec. 3, 2023: Presented Maximal Speaker Separability via Contrastive Learning with Angular Margin and Class-Aware Attention for Hard Samples at International Doctoral Forum 2023, Hong Kong SAR. 🇭🇰
  • May 15, 2023: Paper Sharing Session: Discriminative Speaker Representation via Contrastive Learning with Class-Aware Attention in Angular Space (ICASSP 2023).
  • Jul. 1, 2022: Participant Talk: Shared on speaker verification at Odyssey-CNSRC Workshop 2022.
  • May 29, 2021: 🎓 Completed Master’s oral examination.
  • Nov. 14, 2020: 🏅 CAAI Award: Received the Outstanding Scientific and Technological Achievement Award from the Chinese Association for Artificial Intelligence.
  • Oct. 29, 2020: Video: Uploaded the CCL 2020 oral presentation.
  • Oct. 11, 2020: Video: Uploaded the CCMT 2020 oral presentation.

Research Interests

  • Speech large language models: multilingual modeling, parameter-efficient fine-tuning, and post-training alignment
  • Speech processing: speaker representation learning, speaker verification, and robust speech modeling
  • AI for health: speech-related health applications and multimodal representation learning

Academic Positions

Education

  • PhD in Electrical and Electronic Engineering, The Hong Kong Polytechnic University, Hong Kong SAR — Jan. 2022–Oct. 2025
  • Master in Software Engineering, Xinjiang University, Xinjiang, China — Sept. 2018–June 2021
  • Undergraduate in Computer Science, Qilu University of Technology, Shandong, China — Sept. 2012–June 2016

Selected Awards

  • Best Student Paper Runner-Up Award, Pacific Rim International Conference on Artificial Intelligence (PRICAI 2024), November 2024
  • Outstanding Scientific and Technological Achievement Award, Chinese Association for Artificial Intelligence (CAAI), October 2020
  • Top 100 Teams Award, Intel Cup’s Inaugural Chinese Graduate Student Artificial Intelligence Innovation Competition, June 2019