Ph.D. Candidate, NTU EECS · Taipei, Taiwan
I am a Ph.D. student at National Taiwan University (NTU), advised by Jyh-Shing Roger Jang and Hung-yi Lee.
Prior to this, I earned my M.S. and B.S. in Computer Science from
NTU (2023) and
Taiwan Tech (2020).
My research focuses on large language models, audio language models, audio generation, and deepfakes, with more than 10 first/co-first papers published in top-tier venues such as TASLP, TIST, COLM, ICASSP, INTERSPEECH, ASRU, and SLT.
I have been fortunate to receive several research honors, highlighted by:
IEEE Signal Processing Society Scholarship.Recently, I have been working on training and benchmarking models for audio understanding and generation, building on my earlier work across three main threads: leading research on retrieval and agentic LLMs (MetaBench-Harness, CodaRAG, GDP-RAG); making core contributions to large audio-language models and benchmarks (DeSTA 2.5-Audio, Dynamic-SUPERB Phase-2, Codec-SUPERB); and spearheading work on audio deepfakes (SingGraph, CodecFake+, SASTNet). I am always glad to explore potential collaborations with those who share similar research interests, and I am actively looking for a research scientist internship position.
Feel free to reach out to me at d12942018 [at] ntu.edu.tw, or find me on LinkedIn, Scholar, GitHub, 𝕏.