Ph.D. Candidate, NTU EECS · Taipei, Taiwan
I am a Ph.D. student at National Taiwan University (NTU), advised by Jyh-Shing Roger Jang and Hung-yi Lee.
Prior to this, I earned my M.S. and B.S. in Computer Science from
NTU (2023) and
Taiwan Tech (2020).
My research interests include large language models, spoken language models, speech generation, and deepfakes.
Along the way, I have been fortunate to receive several research honors, highlighted by:
IEEE Signal Processing Society Scholarship.Currently, I am exploring model training and benchmarking for audio understanding and generation. Previously, my research spanned three threads: leading work on audio deepfakes (SingGraph, CodecFake+, SASTNet), contributing to large audio-language models and benchmarks (DeSTA 2.5-Audio, Dynamic-SUPERB Phase-2, Codec-SUPERB), and mentoring research on LLMs for RAG efficiency and their inherent limits.
Feel free to reach out to me at d12942018 [at] ntu.edu.tw, or find me on LinkedIn, Scholar, GitHub, 𝕏.