I am a final-year Ph.D. student at Johns Hopkins University, fortunate to be advised by Prof. Mounya Elhilali.
My Research focuses on AI for understanding and generating sound. My work spans topics such as controllable generation, voice design, source separation, and detection. I enjoy bringing research to life through interactive demos, open-source tools, and datasets for the broader audio and machine learning community.
Outside of academia, I make Music as a producer, mostly hip-hop and hyperpop. This experience shapes how I approach audio research, giving me a creator's perspective on what makes audio tools expressive, controllable, and genuinely useful in the creative process.
At Tsinghua, I studied civil engineering and information management, taking courses such as fluid mechanics and econometrics while exploring data-driven research in business analytics and medical informatics. This quantitative foundation later came together with my long-standing interest in music, leading me to audio research.
Seeking Full-Time Positions
I am open to full-time Research Scientist or
Research Engineer positions, with a focus on
audio, speech, or music technology, starting in early 2027.
Please feel free to contact me at jhai2@jhu.edu if there is a potential fit.