Portrait of Samuele Cornell

Samuele Cornell

Postdoctoral Research Associate
WAVLab · Language Technologies Institute · Carnegie Mellon University

I work on robust “speech-in-the-wild” processing: overlapping speakers, noise, reverberation, and distant microphones. My research spans the full speech pipeline — separation, enhancement, speaker diarization, and automatic speech recognition — with a focus on end-to-end integration and foundation models that natively understand multi-talker conversational speech. I also work on text-to-speech and speech generation for synthetic data creation, enabling speech-in-the-wild systems in domains where real data is scarce, such as doctor–patient conversations.

I am currently a postdoctoral researcher in Prof. Shinji Watanabe's WAVLab at Carnegie Mellon University. I received my PhD in Information Engineering (2023) and MSc in Electronic Engineering (summa cum laude, 2019) from Università Politecnica delle Marche, advised by Prof. Stefano Squartini. Along the way I have worked with Amazon Alexa (applied scientist intern), consulted for startups building AI-powered hearing aids (Fortell, formerly Chromatic, which has since raised $163M at a $740M valuation) and silent speech interfaces (AlterEgo), and took part in three JHU JSALT workshops — as a participant in 2019 (during my master's) and 2026, and leading, with Lukáš Burget, the JSALT 2025 EMMA team on end-to-end multi-channel multi-talker ASR, which has produced six papers to date.

I believe strongly in open and reproducible science: I co-authored and contribute to several widely used open-source speech toolkits (SpeechBrain, ESPnet, Asteroid, TorchAudio — collectively exceeding 23,000 GitHub stars), and I organize community challenges — CHiME-7/8 DASR, URGENT, DCASE Task 4 — that push the field toward robustness and generalization. Core architectures I co-developed, such as SepFormer and TF-GridNet, are widely adopted by industrial research labs — including Google, Microsoft, MERL, and Samsung — for speech separation and enhancement. I am an elected member of the IEEE Speech and Language Processing Technical Committee (SLTC, Speech Recognition area, 2027–2029), served as secretary of the ISCA Special Interest Group on Robust Speech Processing (2023–2025), and routinely serve as meta-reviewer for ICASSP and Interspeech, with 60+ publications and 4,000+ citations in these areas.

When I am not doing research, I like to hike, ski (picture above), swim or — very recently — play tennis and disc golf (I am bad at both; I used to be decent at soccer though). I used to play the electric guitar; now I only play the keyboard (only the laptop one, sadly).

News

Open-Source Software

Selected honors

Press

Teaching

Publications

Selected from 60+ publications — see Google Scholar for the full list.

Most cited

Loading…

Recent

Loading…

Contact

samuele.cornell [at] ieee.org
Language Technologies Institute, Carnegie Mellon University
5000 Forbes Avenue, Pittsburgh, PA

Always happy to chat about robust speech processing, open-source, challenges, or potential collaborations.