Sid.
PhD · Soundability Lab · University of Michigan

Sidharth

Teaching machines to hear the one voice you're talking to.

Speech & audio AI + HCI. I build target speech extraction, speaker representations and real-time enhancement models that work in the noisy rooms where conversations actually happen. Advised by Dhruv Jain, with Hao-Wen Dong.

↳ move over the room: the voice nearest your cursor is extracted

Sidharth at Emerald Bay, Lake Tahoe
SUMMER 2026Spatial audio ML intern
at Bose Research
Research

Three threads, one ear

Robust speech and audio perception in real-world conditions, from the model to the person wearing it.

Enrollment-free target speech extraction

Pulling out the person you're engaging with — without asking them for a reference clip first — so the system stays out of the way of the conversation.

12.83 dB SI-SDRi · ongoing

Speaker representations from mixtures

Learning persistent speaker identities directly from overlapped, multi-talker audio instead of clean single-speaker recordings.

Real-time enhancement & sound control

Low-latency models on raw waveforms, and systems that let people turn down the sounds that overwhelm them.

Publications

Papers

  1. 2026

    First authorUnmixing the Crowd: Learning Persistent Speaker Representations from Mixture-Derived Multi-Speaker Embeddings

    FNU Sidharth, Meysam Asgari, Hao-Wen Dong, Dhruv Jain

    In submission
  2. 2026

    Sona: Real-Time Multi-Target Sound Attenuation for Noise Sensitivity

    Jeremy Zhengqi Huang, Emani Hicks, Sidharth, Gillian R. Hayes, Dhruv Jain

    CHI EA 2026
  3. 2025

    aTENNuate: Optimized Real-time Speech Enhancement with Deep SSMs on Raw Audio

    Yan Ru Pei, Ritik Shrivastava, FNU Sidharth

    Interspeech 2025
  4. 2025

    First authorPainDECOG: Machine Learning-Based Identification of Pain Biomarkers from sEEG Signals

    Sidharth, Vishwas Sathish, Shweta Bansal, Samantha Sun, Timmy Pham, Kurt Weaver, Rajesh P. N. Rao, Jeffrey Herron

    AAAI 2025 · W3PHIAI
  5. 2023

    The DISPLACE Challenge 2023: DIarization of SPeaker and LAnguage in Conversational Environments

    Shikha Baghel, Shreyas Ramoji, Sidharth, et al.

    Interspeech 2023
  6. 2023

    First authorEmotion Detection from EEG using Transfer Learning

    Sidharth, et al.

    IEEE EMBC 2023
  7. 2023

    CSP-LSTM Based Emotion Recognition from EEG Signals

    Jerrin Thomas Panachakel, H. Ranjana, Sidharth, et al.

    IEEE MetroXRAINE 2023
  8. 2023

    EEG-based Emotion Classification: A Theoretical Perusal of Deep Learning Methods

    K. Sana Parveen, Jerrin Thomas Panachakel, H. Ranjana, Sidharth, et al.

    IEEE INOCON 2023
News

Lately

Submitted Unmixing the Crowd to IEEE SLT 2026.

Joined Bose Research for the summer — open-vocabulary, queryable sound event localization.

Sona appears at CHI 2026 Extended Abstracts.

Started the PhD in CSE at the University of Michigan.

aTENNuate accepted at Interspeech 2025.

Presented PainDECOG at AAAI 2025 W3PHIAI.

Experience

Industry & education

IndustryAudio ML research at product companies
  1. Bose logo
    Jun — Aug 2026Bose ResearchResearch Intern · Spatial Audio / Machine Learning
  2. Skyworks logo
    May — Aug 2025Skyworks SolutionsResearch Intern
  3. BrainChip logo
    Jun 2024 — Mar 2025BrainChip ResearchResearch Intern
EducationInstrumentation → electrical engineering → computer science
  1. University of Michigan CSE logo
    2025 — 2029 (expected) PhD, Computer Science & Engineering University of Michigan, Ann Arbor Advisor Dhruv Jain · collaborator Hao-Wen Dong Soundability Lab logoSoundability Lab
  2. University of Washington logo
    2023 — 2025 MS, Electrical Engineering University of Washington, Seattle Advisors Rajesh Rao & Jeffrey Herron Thesis · Decoding Pain: Statistical Identification of Biomarkers from Electrophysiological Signals
  3. College of Engineering Trivandrum crest
    2019 — 2023 B.Tech, Electronics & Instrumentation College of Engineering Trivandrum · minor in Mathematics Advisor Jerrin Thomas Panachakel Thesis · Emotion Detection from EEG using Transfer Learning
Talks

On stage

Decoding Pain: Statistical Identification of Biomarkers from Electrophysiological SignalsAAAI 2025 · W3PHIAI workshop · Mar 2025
Emotion Detection from EEG using Transfer LearningIEEE EMBC 2023 · Sydney · Jul 2023
Contact

Let's talk about voices in noise.

sidcs@umich.edu