Jiarui Hai
Jiarui Hai
Making machines understand and generate sound
Education
  • Johns Hopkins University
    Johns Hopkins University
    Aug. 2022 - Present
    PhD Candidate @ ECE & DSAI
    MD, USA
  • Tsinghua University
    Tsinghua University
    Aug. 2020 – Jun. 2022
    Master of Engineering
    Beijing, China
  • Tsinghua University
    Tsinghua University
    Aug. 2016 – Jun. 2020
    Bachelor of Engineering
    Beijing, China
    Bachelor of Science
  • Experience
  • Apple
    Apple
    Apr 2026 - Present
    ML Research Intern
    USA
  • Adobe | CAVA
    Adobe | CAVA
    May 2025 - Feb. 2026
    Research Scientist Intern
    CA, USA
  • Tencent | AI Lab
    Tencent | AI Lab
    May 2024 - Sep. 2024
    Research Scientist Intern
    WA, USA
  • Kuaishou | AI Platform
    Kuaishou | AI Platform
    Aug. 2021 - Feb. 2022
    Music Technology Intern
    Beijing, China
  • About Me
    Introducing OpenSound
    OpenSound is a community-driven initiative led by researchers from Johns Hopkins University, dedicated to advancing research in audio, speech, and music. The project brings together contributors to explore and develop interactive demos, build and refine models, curate high-quality datasets, and establish meaningful benchmarks for evaluation.
    News
    FlexSED was selected as a Spotlight at WASPAA.
    FlexSED was selected as a Spotlight at WASPAA. Oct 2025
    Oct 2025
    EzAudio was accepted as an oral presentation at Interspeech.
    EzAudio was accepted as an oral presentation at Interspeech. Jun 2025
    Jun 2025
    We built OpenSound to share speech and audio models with live demos.
    We built OpenSound to share speech and audio models with live demos. Sep 2024
    Sep 2024
    Gave an oral presentation of my first first-author paper, Diff-Pitcher, at WASPAA.
    Gave an oral presentation of my first first-author paper, Diff-Pitcher, at WASPAA. Oct 2023
    Oct 2023
    Joined LCAP @ Johns Hopkins University as a PhD student.
    Joined LCAP @ Johns Hopkins University as a PhD student. Aug 2022
    Aug 2022
    * Equal contribution,
    Research mentorship
    Research Highlights
    CapSpeech: Enabling Downstream Applications in Style-Captioned Text-to-Speech
    CapSpeech: Enabling Downstream Applications in Style-Captioned Text-to-Speech

    Helin Wang*, Jiarui Hai*, Dading Chong, Karan Thakkar, Tiantian Feng, Dongchao Yang

    TASLP, 2026

    A style-captioned TTS dataset and framework enabling controllable, style-aware text-to-speech for downstream applications.

    FlexSED: Towards Open-Vocabulary Sound Event Detection
    FlexSED: Towards Open-Vocabulary Sound Event Detection

    Jiarui Hai, Helin Wang, Weizhe Guo, Mounya Elhilali

    WASPAA, 2025 Spotlight

    An open-vocabulary sound event detection approach that generalizes to unseen classes via flexible text-conditioned modeling.

    EzAudio: Enhancing Text-to-Audio Generation with Efficient Diffusion Transformer
    EzAudio: Enhancing Text-to-Audio Generation with Efficient Diffusion Transformer

    Jiarui Hai, Yong Xu, Hao Zhang, Chenxing Li, Helin Wang, Mounya Elhilali

    Interspeech, 2025 Oral

    An efficient diffusion transformer that improves text-to-audio generation quality while reducing compute.

    Diff-Pitcher: Diffusion-based Singing Voice Pitch Correction
    Diff-Pitcher: Diffusion-based Singing Voice Pitch Correction

    Jiarui Hai, Mounya Elhilali

    WASPAA, 2023 Oral

    A diffusion-based method for singing voice pitch correction that adjusts pitch while preserving timbre and expression.

    More Projects
    Education
  • Johns Hopkins University
    Johns Hopkins University
    Aug. 2022 - Present
    PhD Candidate @ ECE & DSAI
    MD, USA
  • Tsinghua University
    Tsinghua University
    Aug. 2020 – Jun. 2022
    Master of Engineering
    Beijing, China
  • Tsinghua University
    Tsinghua University
    Aug. 2016 – Jun. 2020
    Bachelor of Engineering
    Beijing, China
    Bachelor of Science
  • Experience
  • Apple
    Apple
    Apr 2026 - Present
    ML Research Intern
    USA
  • Adobe | CAVA
    Adobe | CAVA
    May 2025 - Feb. 2026
    Research Scientist Intern
    CA, USA
  • Tencent | AI Lab
    Tencent | AI Lab
    May 2024 - Sep. 2024
    Research Scientist Intern
    WA, USA
  • Kuaishou | AI Platform
    Kuaishou | AI Platform
    Aug. 2021 - Feb. 2022
    Music Technology Intern
    Beijing, China