vix.ing · top · new · best · stats · spec

Manuel Sam Ribeiro

  1. TaL: a synchronised multi-speaker corpus of ultrasound tongue imaging,\n audio, and lip videos
    2020/11/19 by Manuel Sam Ribeiro, Ribeiro, Manuel Sam, Jennifer Sanger +11 · 4 citations
    Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Image and Video Processing (eess.IV) #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  2. Voice Filter: Few-shot text-to-speech speaker adaptation using voice conversion as a post-processing module
    2022/02/16 by Adam Gabryś, Goeric Huybrechts, Gabryś, Adam +15 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
  3. Improving grapheme-to-phoneme conversion by learning pronunciations from speech recordings
    2023/07/31 by Manuel Sam Ribeiro, Ribeiro, Manuel Sam, Giulia Comini +3 · 1 citation
    Computer Science · #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling
  4. Cross-speaker style transfer for text-to-speech using data augmentation
    2022/02/10 by Manuel Sam Ribeiro, Julian Roth, Ribeiro, Manuel Sam +9 · 1 citation
    Computer Science · #Advanced Data Compression Techniques #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  5. Low-data? No problem: low-resource, language-agnostic conversational text-to-speech via F0-conditioned data augmentation
    2022/07/29 by Giulia Comini, Goeric Huybrechts, Comini, Giulia +7 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Topic Modeling #electronic engineering #information engineering