Alessio Brutti
- MOSEL: 950,000 Hours of Speech Data for Open-Source Speech Foundation Model Training on EU Languages
2024/10/01 by Marco Gaido, Gaido, Marco, Sara Papi +15 · 5 citations
Computer Science · #Natural Language Processing Techniques #Speech Recognition and Synthesis
- Low-Latency Speech Separation Guided Diarization for Telephone Conversations
2022/04/05 by Giovanni Morrone, Morrone, Giovanni, Samuele Cornell +10 · 1 citation
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Phonetics and Phonology Research #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Parameter-Efficient Transfer Learning of Audio Spectrogram Transformers
2023/12/06 by Umberto Cappellazzo, Daniele Falavigna, Cappellazzo, Umberto +5 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Scaling and Enhancing LLM-based AVSR: A Sparse Mixture of Projectors Approach
2025/05/20 by Umberto Cappellazzo, Minsu Kim, Cappellazzo, Umberto +7 · 3 citations
Computer Science · #Speech and Audio Processing #Speech Recognition and Synthesis #Music and Audio Processing
- MEUSLI: a Multilingual Projector for LLM-based ASR and Beyond
2026/07/24 by Lorenzo Concina, Seraphina Fong, Marco Matassoni +1
#cs.CL #cs.AI #eess.AS
- SpeechLLM Meets Federated Learning for End-to-End ASR: English and Italian Case Studies
2026/07/28 by Mohamed Nabih Ali, Daniele Falavigna, Alessio Brutti
#cs.CL