Edresson Casanova
- MLAAD: The Multi-Language Audio Anti-Spoofing Dataset
2024/01/17 by Nicolas M. Müller, Müller, Nicolas M., Piotr Kawa +15 · 44 citations
Computer Science · Medicine · #Speech Recognition and Synthesis #Voice and Speech Disorders #Music and Audio Processing
- BibleTTS: a large, high-fidelity, multilingual, and uniquely African speech corpus
2022/07/07 by Josh Meyer, Meyer, Josh, David Ifeoluwa Adelani +35 · 8 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Koel-TTS: Enhancing LLM based Speech Generation with Preference Alignment and Classifier Free Guidance
2025/02/07 by Shehzeen Hussain, Paarth Neekhara, Hussain, Shehzeen +15 · 18 citations
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- ASR data augmentation in low-resource settings using cross-lingual multi-speaker TTS and cross-lingual voice conversion
2022/03/29 by Edresson Casanova, Casanova, Edresson, Christopher Shulby +10 · 4 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
- CML-TTS A Multilingual Dataset for Speech Synthesis in Low-Resource Languages
2023/06/16 by Frederico Santos de Oliveira, Oliveira, Frederico S., Edresson Casanova +6 · 5 citations
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- SALM-Duplex: Efficient and Direct Duplex Modeling for Speech-to-Speech Language Model
2025/05/21 by Ke Hu, Hu, Ke, Ehsan Hosseini-Asl +17 · 13 citations
Computer Science · #Speech Recognition and Synthesis #Speech and dialogue systems #Speech and Audio Processing
- Low Frame-rate Speech Codec: a Codec Designed for Fast High-quality Speech LLM Training and Inference
2024/09/18 by Edresson Casanova, Ryan Langman, Casanova, Edresson +13 · 9 citations
Computer Science · #Advanced Data Compression Techniques #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- CORAA: a large corpus of spontaneous and prepared speech manually validated for speech recognition in Brazilian Portuguese
2021/10/14 by Junior, Arnaldo Candido, Edresson Casanova, Casanova, Edresson +18 · 2 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Brazilian Portuguese Speech Recognition Using Wav2vec 2.0
2021/07/23 by Lucas Rafael Stefanel Gris, Edresson Casanova, Gris, Lucas Rafael Stefanel +6 · 1 citation
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Music and Audio Processing #Natural Language Processing Techniques #Speech Recognition and Synthesis
- HiFiTTS-2: A Large-Scale High Bandwidth Speech Dataset
2025/06/04 by Ryan Langman, Langman, Ryan, Xuesong Yang +11 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- NanoCodec: Towards High-Quality Ultra Fast Speech LLM Inference
2025/08/07 by Edresson Casanova, Paarth Neekhara, Casanova, Edresson +15 · 5 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering