vix.ing · top · new · best · stats · spec

Edresson Casanova

  1. MLAAD: The Multi-Language Audio Anti-Spoofing Dataset
    2024/01/17 by Nicolas M. Müller, Müller, Nicolas M., Piotr Kawa +15 · 44 citations
    Computer Science · Medicine · #Speech Recognition and Synthesis #Voice and Speech Disorders #Music and Audio Processing
  2. BibleTTS: a large, high-fidelity, multilingual, and uniquely African speech corpus
    2022/07/07 by Josh Meyer, Meyer, Josh, David Ifeoluwa Adelani +35 · 8 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  3. Koel-TTS: Enhancing LLM based Speech Generation with Preference Alignment and Classifier Free Guidance
    2025/02/07 by Shehzeen Hussain, Paarth Neekhara, Hussain, Shehzeen +15 · 18 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
  4. ASR data augmentation in low-resource settings using cross-lingual multi-speaker TTS and cross-lingual voice conversion
    2022/03/29 by Edresson Casanova, Casanova, Edresson, Christopher Shulby +10 · 4 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
  5. CML-TTS A Multilingual Dataset for Speech Synthesis in Low-Resource Languages
    2023/06/16 by Frederico Santos de Oliveira, Oliveira, Frederico S., Edresson Casanova +6 · 5 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
  6. SALM-Duplex: Efficient and Direct Duplex Modeling for Speech-to-Speech Language Model
    2025/05/21 by Ke Hu, Hu, Ke, Ehsan Hosseini-Asl +17 · 13 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and dialogue systems #Speech and Audio Processing
  7. Low Frame-rate Speech Codec: a Codec Designed for Fast High-quality Speech LLM Training and Inference
    2024/09/18 by Edresson Casanova, Ryan Langman, Casanova, Edresson +13 · 9 citations
    Computer Science · #Advanced Data Compression Techniques #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  8. CORAA: a large corpus of spontaneous and prepared speech manually validated for speech recognition in Brazilian Portuguese
    2021/10/14 by Junior, Arnaldo Candido, Edresson Casanova, Casanova, Edresson +18 · 2 citations
    Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  9. Brazilian Portuguese Speech Recognition Using Wav2vec 2.0
    2021/07/23 by Lucas Rafael Stefanel Gris, Edresson Casanova, Gris, Lucas Rafael Stefanel +6 · 1 citation
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Music and Audio Processing #Natural Language Processing Techniques #Speech Recognition and Synthesis
  10. HiFiTTS-2: A Large-Scale High Bandwidth Speech Dataset
    2025/06/04 by Ryan Langman, Langman, Ryan, Xuesong Yang +11 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  11. NanoCodec: Towards High-Quality Ultra Fast Speech LLM Inference
    2025/08/07 by Edresson Casanova, Paarth Neekhara, Casanova, Edresson +15 · 5 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering