Salah Zaiem
- Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
2025/07/07 by Gheorghe Comanici, Comanici, Gheorghe, Eric Bieber +6844 · 8 voices · 1467 citations
Computer Science · #cs.CL #cs.AI
- Open-Source Conversational AI with SpeechBrain 1.0
2024/06/29 by Mirco Ravanelli, Titouan Parcollet, Ravanelli, Mirco +60 · 39 citations
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Human-Computer Interaction (cs.HC) #Machine Learning (cs.LG) #Natural Language Processing Techniques #Speech and dialogue systems #electronic engineering #information engineering
- How Should We Extract Discrete Audio Tokens from Self-Supervised Models?
2024/06/15 by Pooneh Mousavi, Mousavi, Pooneh, Jarod Duret +11 · 13 citations
Arts and Humanities · Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #Diverse Musicological Studies #FOS: Computer and information sciences #FOS: Electrical engineering #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #electronic engineering #information engineering
- DP-Parse: Finding Word Boundaries from Raw Speech with an Instance Lexicon
2022/06/22 by Robin Algayres, Tristan Ricoul, Algayres, Robin +13 · 4 citations
Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Topic Modeling
- End-to-End Speech Recognition from Federated Acoustic Models
2021/04/29 by Yan Gao, Titouan Parcollet, Gao, Yan +11 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Fine-tuning Strategies for Faster Inference using Speech Self-Supervised Models: A Comparative Study
2023/03/12 by Salah Zaiem, Robin Algayres, Zaiem, Salah +7 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Training dynamic models using early exits for automatic speech recognition on resource-constrained devices
2023/09/18 by George August Wright, Wright, George August, Umberto Cappellazzo +11 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- Less Forgetting for Better Generalization: Exploring Continual-learning Fine-tuning Methods for Speech Self-supervised Representations
2024/06/30 by Salah Zaiem, Titouan Parcollet, Zaiem, Salah +3 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering