vix.ing · top · new · best · stats · spec

Alessio Brutti

  1. MOSEL: 950,000 Hours of Speech Data for Open-Source Speech Foundation Model Training on EU Languages
    2024/10/01 by Marco Gaido, Gaido, Marco, Sara Papi +15 · 5 citations
    Computer Science · #Natural Language Processing Techniques #Speech Recognition and Synthesis
  2. Low-Latency Speech Separation Guided Diarization for Telephone Conversations
    2022/04/05 by Giovanni Morrone, Morrone, Giovanni, Samuele Cornell +10 · 1 citation
    Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Phonetics and Phonology Research #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  3. Parameter-Efficient Transfer Learning of Audio Spectrogram Transformers
    2023/12/06 by Umberto Cappellazzo, Daniele Falavigna, Cappellazzo, Umberto +5 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  4. Scaling and Enhancing LLM-based AVSR: A Sparse Mixture of Projectors Approach
    2025/05/20 by Umberto Cappellazzo, Minsu Kim, Cappellazzo, Umberto +7 · 3 citations
    Computer Science · #Speech and Audio Processing #Speech Recognition and Synthesis #Music and Audio Processing
  5. MEUSLI: a Multilingual Projector for LLM-based ASR and Beyond
    2026/07/24 by Lorenzo Concina, Seraphina Fong, Marco Matassoni +1
    #cs.CL #cs.AI #eess.AS
  6. SpeechLLM Meets Federated Learning for End-to-End ASR: English and Italian Case Studies
    2026/07/28 by Mohamed Nabih Ali, Daniele Falavigna, Alessio Brutti
    #cs.CL