Lorenzo-Trueba, Jaime
- The Voice Conversion Challenge 2018: Promoting Development of Parallel and Nonparallel Methods
2018/04/12 by Jaime Lorenzo-Trueba, Junichi Yamagishi, Lorenzo-Trueba, Jaime +11 · 16 citations
Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (stat.ML) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Voice and Speech Disorders #electronic engineering #information engineering
- EmoCat: Language-agnostic Emotional Voice Conversion
2021/01/14 by Schnell, Bastian, Huybrechts, Goeric, Perz, Bartek +2 · 2 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Weakly-supervised word-level pronunciation error detection in non-native\n English speech
2021/06/07 by Daniel Korzekwa, Jaime Lorenzo-Trueba, Korzekwa, Daniel +7 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- Low-resource expressive text-to-speech using data augmentation
2020/11/11 by Huybrechts, Goeric, Merritt, Thomas, Comini, Giulia +3 · 2 citations
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- A comparison of recent waveform generation and acoustic modeling methods for neural-network-based speech synthesis
2018/04/07 by Wang, Xin, Lorenzo-Trueba, Jaime, Takaki, Shinji +2 · 1 citation
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (stat.ML) #Sound (cs.SD) #electronic engineering #information engineering
- Enhancing the Stability of LLM-based Speech Generation Systems through Self-Supervised Representations
2024/02/05 by Álvaro Martín-Cortinas, Daniel Sáez-Trigueros, Martín-Cortinas, Álvaro +15 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- Voicy: Zero-Shot Non-Parallel Voice Conversion in Noisy Reverberant\n Environments
2021/06/16 by Alejandro Mottini, Mottini, Alejandro, Jaime Lorenzo-Trueba +5 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Mispronunciation Detection in Non-native (L2) English with Uncertainty Modeling
2021/01/16 by Korzekwa, Daniel, Lorenzo-Trueba, Jaime, Zaporowski, Szymon +3 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- Voice Filter: Few-shot text-to-speech speaker adaptation using voice conversion as a post-processing module
2022/02/16 by Adam Gabryś, Goeric Huybrechts, Gabryś, Adam +15 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
- Improving grapheme-to-phoneme conversion by learning pronunciations from speech recordings
2023/07/31 by Manuel Sam Ribeiro, Giulia Comini, Ribeiro, Manuel Sam +3 · 1 citation
Computer Science · #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling
- Cross-speaker style transfer for text-to-speech using data augmentation
2022/02/10 by Manuel Sam Ribeiro, Julian Roth, Ribeiro, Manuel Sam +9 · 1 citation
Computer Science · #Advanced Data Compression Techniques #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Deep Encoder-Decoder Models for Unsupervised Learning of Controllable Speech Synthesis
2018/07/30 by Gustav Eje Henter, Henter, Gustav Eje, Jaime Lorenzo-Trueba +5 · 1 citation
Computer Science · #62F99 #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #G.3 #I.2.7 #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Low-data? No problem: low-resource, language-agnostic conversational text-to-speech via F0-conditioned data augmentation
2022/07/29 by Giulia Comini, Goeric Huybrechts, Comini, Giulia +7 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Topic Modeling #electronic engineering #information engineering