Portrait of Samuele Cornell

Samuele Cornell

Postdoctoral Research Associate
WAVLab · Language Technologies Institute · Carnegie Mellon University

I work on robust “speech-in-the-wild” processing: overlapping speakers, noise, reverberation, and distant microphones. My research spans the full speech pipeline — separation, enhancement, speaker diarization, and automatic speech recognition — with a focus on end-to-end integration and foundation models that natively understand multi-talker conversational speech. I also work on text-to-speech and speech generation for synthetic data creation, enabling speech-in-the-wild systems in domains where real data is scarce, such as doctor–patient conversations.

I am currently a postdoctoral researcher in Prof. Shinji Watanabe's WAVLab at Carnegie Mellon University. I received my PhD in Information Engineering (2023) and MSc in Electronic Engineering (summa cum laude, 2019) from Università Politecnica delle Marche, advised by Prof. Stefano Squartini. Along the way I have worked with Amazon Alexa (applied scientist intern), consulted for startups building AI-powered hearing aids (Fortell, formerly Chromatic, which has since raised $163M at a $740M valuation) and silent speech interfaces (AlterEgo), and took part in three JHU JSALT workshops — as a participant in 2019 (during my master's) and 2026, and leading, with Lukáš Burget, the JSALT 2025 EMMA team on end-to-end multi-channel multi-talker ASR, which has produced six papers to date.

I believe strongly in open and reproducible science: I co-authored and contribute to several widely used open-source speech toolkits (SpeechBrain, ESPnet, Asteroid, TorchAudio — collectively exceeding 23,000 GitHub stars), and I organize community challenges — CHiME-7/8 DASR, URGENT, DCASE Task 4 — that push the field toward robustness and generalization. Core architectures I co-developed, such as SepFormer and TF-GridNet, are widely adopted by industrial research labs — including Google, Microsoft, MERL, and Samsung — for speech separation and enhancement. I served as secretary of the ISCA Special Interest Group on Robust Speech Processing (2023–2025) and serve as Area Chair / meta-reviewer for ICASSP 2026, Interspeech 2026, and IJCNN 2025–26, with 60+ publications and 4,000+ citations in these areas.

When I am not doing research, I like to hike, ski (picture above), swim or — very recently — play tennis and disc golf (I am bad at both; I used to be decent at soccer though). I used to play the electric guitar; now I only play the keyboard (only the laptop one, sadly).

News

Open-Source Software

Selected honors

Press

Publications

Selected from 60+ publications — see Google Scholar for the full list.

Most cited

Loading…

Recent

Loading…

Contact

samuele.cornell [at] ieee.org
Language Technologies Institute, Carnegie Mellon University
5000 Forbes Avenue, Pittsburgh, PA

Always happy to chat about robust speech processing, open-source, challenges, or potential collaborations.