My Research focuses on AI for audio, speech, and music. As a Ph.D. student at Johns Hopkins University, I develop systems that enable machines to understand, reason about, and generate sound. I also contribute open-source projects and datasets to the broader audio and speech research community.
Outside of academia, I make Music as an active producer with a deep interest in the intersection of AI and music. I enjoy exploring how emerging technologies expand the creative process and unlock new possibilities for music creation and production.

A style-captioned TTS dataset and framework enabling controllable, style-aware text-to-speech for downstream applications.
Dataset IEEE Transactions on Audio, Speech and Language Processing (TASLP) | 2026
A style-captioned TTS dataset and framework enabling controllable, style-aware text-to-speech for downstream applications.
Dataset IEEE Transactions on Audio, Speech and Language Processing (TASLP) | 2026

An open-vocabulary sound event detection approach that generalizes to unseen classes via flexible text-conditioned modeling.
Spotlight IEEE Workshop on Applications of Signal Processing to Audio and Acoustics (WASPAA) | 2025
An open-vocabulary sound event detection approach that generalizes to unseen classes via flexible text-conditioned modeling.
Spotlight IEEE Workshop on Applications of Signal Processing to Audio and Acoustics (WASPAA) | 2025

An efficient diffusion transformer that improves text-to-audio generation quality while reducing compute.
Oral Interspeech | 2025
An efficient diffusion transformer that improves text-to-audio generation quality while reducing compute.
Oral Interspeech | 2025

A diffusion-based method for singing voice pitch correction that adjusts pitch while preserving timbre and expression.
Oral IEEE Workshop on Applications of Signal Processing to Audio and Acoustics (WASPAA) | 2023
A diffusion-based method for singing voice pitch correction that adjusts pitch while preserving timbre and expression.
Oral IEEE Workshop on Applications of Signal Processing to Audio and Acoustics (WASPAA) | 2023