AudioMOS Challenge · ASRU 2025
- 1st Place Track 3: MOS prediction for speech in high sampling frequencies Paper: HighRateMOS
Ph.D. Student · EECS · National Taiwan University
I’m a second-year Ph.D. student at National Taiwan University, co-advised by Prof. Hung-yi Lee at the Speech Processing & Machine Learning Lab, NTU, and Prof. Yu Tsao at the Bio-ASP Lab, Academia Sinica.
My research interests lie in speech and audio processing, with a focus on speech enhancement and speech quality assessment. My enhancement work spans single-channel, multi-channel, and multimodal settings, and I organized the first Real-World Audio-Visual Speech Enhancement Challenge at ISCSLP 2026. On the assessment side, I won 1st place in the AudioMOS Challenge 2025 Track 3.
I have also contributed to large audio-language models and benchmarks, including Dynamic-SUPERB Phase-2, DeSTA2, and Game-Time.
I am currently working on speech representations for large audio models, neural audio codecs, and LM-based front-end speech processing.
Competitions & awards
05 · Contact
If there’s anything you’d like to chat about, or you’re interested in collaborating, or even if you’d just like to make a new friend, please feel free to reach out anytime.
Scan the QR code to add me on WeChat