Junior, Arnaldo Candido
- SC-GlowTTS: an Efficient Zero-Shot Multi-Speaker Text-To-Speech Model
2021/04/02 by Casanova, Edresson, Shulby, Christopher, Gölge, Eren +6 · 12 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- ASR data augmentation in low-resource settings using cross-lingual multi-speaker TTS and cross-lingual voice conversion
2022/03/29 by Edresson Casanova, Casanova, Edresson, Christopher Shulby +10 · 4 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
- CORAA: a large corpus of spontaneous and prepared speech manually validated for speech recognition in Brazilian Portuguese
2021/10/14 by Junior, Arnaldo Candido, Edresson Casanova, Casanova, Edresson +18 · 2 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Brazilian Portuguese Speech Recognition Using Wav2vec 2.0
2021/07/23 by Lucas Rafael Stefanel Gris, Gris, Lucas Rafael Stefanel, Edresson Casanova +6 · 1 citation
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Music and Audio Processing #Natural Language Processing Techniques #Speech Recognition and Synthesis
- Evaluating OpenAI's Whisper ASR for Punctuation Prediction and Topic Modeling of life histories of the Museum of the Person
2023/05/23 by Gris, Lucas Rafael Stefanel, Marcacini, Ricardo, Junior, Arnaldo Candido +3 · 1 citation
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
- A Large Dataset of Spontaneous Speech with the Accent Spoken in São Paulo for Automatic Speech Recognition Evaluation
2024/09/10 by Rodrigo Lima, Sidney Evaldo Leal, Lima, Rodrigo +4 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Speech Recognition and Synthesis #electronic engineering #information engineering