Dareen Alharthi
/dɑːriːn ɑlhɑːrθi/ | دارين الحارثي
I am a PhD student at Carnegie Mellon University's Language Technologies Institute, advised by Professor Bhiksha Raj in the Machine Learning for Signal Processing (MLSP) Lab, where I also earned my Master's in Language Technologies. Before CMU, I received my Bachelor's degree in Computer Science from IMSIU University, where I worked with Professor Areeb Alowisheq. My research focuses on voice editing and speech generative models.
News
- [Jun 2026] "RIVET: Robust Idempotent Voice Attribute Editing" accepted at INTERSPEECH 2026.
- [May 2026] "Encoder-Decoder Manifold Alignment for Idempotent Generation" (preprint).
- [Jul 2026] "Heard but Not Heeded: Paralinguistic Information Encoding and Loss in Audio-Language Models" accepted at COLM 2026.
- [Jun 2026] "Subliminal Prosody Learning: Auxiliary Emotion Supervision Redistributes Affective Representations Across ALM Layers" accepted at the Mechanistic Interpretability Workshop, ICML 2026.
- [Dec 2024] "Tessellated Linear Model for Age Prediction from Voice" accepted at IEEE ICASSP 2025.
- [Jul 2024] "Evaluating Speech Synthesis by Training Recognizers on Synthetic Speech" accepted at SynData4GenAI Workshop, INTERSPEECH 2024.
- [Jun 2024] "PAM: Prompting Audio-Language Models for Audio Quality Assessment" accepted at INTERSPEECH 2024.
- [Mar 2025] Gave a guest lecture on "Transformers" for the Introduction to Deep Learning course at CMU.
Watch the lecture