Ph.D. Candidate, NTU EECS · Taipei, Taiwan
I am a Ph.D. candidate at National Taiwan University (NTU), advised by Jyh-Shing Roger Jang and Hung-yi Lee.
My research explores the intersection of large language models and speech/audio processing, focusing primarily on training and benchmarking foundation models for audio understanding and generation.
I have published 10+ first/co-first papers in top-tier venues (e.g., COLM, TASLP, TIST, ICASSP, INTERSPEECH), with selected honors including the NVIDIA Academic Grant, the
IEEE Signal Processing Society Scholarship, and a Best Student Paper nomination at ASRU 2025.
Additionally, I am actively looking for a 2027 research internship.
Feel free to reach out at d12942018 [at] ntu.edu.tw, or find me on LinkedIn, GitHub, and 𝕏.
See all publications by topic, or my Google Scholar for a full list.