Speech Editing & Synthesis
Noise-resilient, zero-shot speech editing with in-context enhancement for insertion and replacement in real-world acoustic conditions.
Speech · Audio · Machine Learning
Ph.D. Student & AI Researcher
I work on speech and audio intelligence at National Taiwan University (GICE) and the AI Research Center, Inventec — spanning zero-shot speech editing, audio-watermark security, bias in instruction TTS, and music information retrieval. I am part of the DISP Lab (NTU GICE) and the Bio-ASP Lab (Academia Sinica).
I am a direct Ph.D. student at National Taiwan University working at the intersection of generative audio and trustworthy AI. My research develops speech systems that remain useful under real-world acoustic conditions and examines how they can be protected, attacked, or biased.
I conduct research with the Bio-ASP Lab at Academia Sinica and the AI Research Center at Inventec, collaborating on robust speech generation, fairness in speech AI, and audio security.
Noise-resilient, zero-shot speech editing with in-context enhancement for insertion and replacement in real-world acoustic conditions.
Watermark-as-trigger backdoors, data-to-model ownership protection, and robustness against acoustic filtering and model pruning.
Compositional analysis of gender bias across social status, career, and persona cues, uncovering multi-dimensional interaction effects.
Query-by-humming, instrument analysis, and diffusion-based guitar tone morphing in latent audio spaces.
* Equal contribution
No publications match this filter.
National Taiwan University
Advisors: Yu Tsao and Jian-Jiun Ding.
National Taiwan University
Research on ECG watermarking for biomedical signals, alongside collaborative work in speech and audio AI.
Inventec Corporation · AI Research Center
Research mentor: Jeng-Lin Li.
Leaderg
I’m always glad to discuss speech and audio research, collaborations, or new ideas — reach out through any channel below.