vix.ing · top · new · best · stats · spec

Junior, Arnaldo Candido

  1. SC-GlowTTS: an Efficient Zero-Shot Multi-Speaker Text-To-Speech Model
    2021/04/02 by Casanova, Edresson, Shulby, Christopher, Gölge, Eren +6 · 12 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  2. ASR data augmentation in low-resource settings using cross-lingual multi-speaker TTS and cross-lingual voice conversion
    2022/03/29 by Edresson Casanova, Casanova, Edresson, Christopher Shulby +10 · 4 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
  3. CORAA: a large corpus of spontaneous and prepared speech manually validated for speech recognition in Brazilian Portuguese
    2021/10/14 by Junior, Arnaldo Candido, Edresson Casanova, Casanova, Edresson +18 · 2 citations
    Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  4. Brazilian Portuguese Speech Recognition Using Wav2vec 2.0
    2021/07/23 by Lucas Rafael Stefanel Gris, Gris, Lucas Rafael Stefanel, Edresson Casanova +6 · 1 citation
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Music and Audio Processing #Natural Language Processing Techniques #Speech Recognition and Synthesis
  5. Evaluating OpenAI's Whisper ASR for Punctuation Prediction and Topic Modeling of life histories of the Museum of the Person
    2023/05/23 by Gris, Lucas Rafael Stefanel, Marcacini, Ricardo, Junior, Arnaldo Candido +3 · 1 citation
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
  6. A Large Dataset of Spontaneous Speech with the Accent Spoken in São Paulo for Automatic Speech Recognition Evaluation
    2024/09/10 by Rodrigo Lima, Sidney Evaldo Leal, Lima, Rodrigo +4 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Speech Recognition and Synthesis #electronic engineering #information engineering