Joly, Arnaud
- BASE TTS: Lessons from building a billion-parameter Text-to-Speech model on 100K hours of data
2024/02/12 by Mateusz Łajszczak, Guillermo Cámbara, Łajszczak, Mateusz +36 · 2 voices · 24 citations
Computer Science · Engineering · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #cs.CL #cs.LG #eess.AS #electronic engineering #information engineering
- Controllable Emphasis with zero data for text-to-speech
2023/07/13 by Arnaud Joly, Marco Nicolis, Joly, Arnaud +25 · 2 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Prosodic Representation Learning and Contextual Sampling for Neural\n Text-to-Speech
2020/11/04 by Sri Karlapati, Ammar N. Abbas, Karlapati, Sri +11 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering