vix.ing · top · new · best · stats
  1. Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers
    2023/01/05 by Chengyi Wang, Wang, Chengyi, Sanyuan Chen +23 · 1 voice · 268 citations
    Computer Science · Engineering · #Audio and Speech Processing (eess.AS) #Code-excited linear prediction #Codec #Codec2 #Computation and Language (cs.CL) #Computer science #Context (archaeology) #FOS: Computer and information sciences #FOS: Electrical engineering #Language model #Linear predictive coding #Natural language processing #Naturalness #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #Speech coding #Speech recognition #Speech synthesis #Task (project management) #Topic Modeling #cs.CL #cs.SD #eess.AS #electronic engineering #information engineering
  2. WaveCycleGAN2: Time-domain Neural Post-filter for Speech Waveform Generation
    2019/04/05 by Kou Tanaka, Tanaka, Kou, Hirokazu Kameoka +5 · 18 citations
    Computer Science · Engineering · Mathematics · #Aliasing #Artificial intelligence #Artificial neural network #Audio and Speech Processing (eess.AS) #Computer science #FOS: Computer and information sciences #FOS: Electrical engineering #Filter (signal processing) #Linear predictive coding #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Mathematics #Music and Audio Processing #Parametric statistics #Sampling (signal processing) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech coding #Speech processing #Speech recognition #Speech synthesis #Statistics #Telecommunications #Time domain #Voice activity detection #Waveform #cs.LG #cs.SD #eess.AS #electronic engineering #information engineering #stat.ML
  3. Speex: A Free Codec For Free Speech
    2016/02/28 by Jean-Marc Valin, Valin, Jean-Marc · 53 citations
    Computer Science · Engineering · #Adaptive Multi-Rate audio codec #Advanced Adaptive Filtering Techniques #Advanced Data Compression Techniques #Code-excited linear prediction #Codec #Codec2 #Computer network #Computer science #FOS: Computer and information sciences #Free speech #Full Rate #Latency (audio) #Linear predictive coding #Network packet #Sound (cs.SD) #Speech and Audio Processing #Speech coding #Speech processing #Speech recognition #Telecommunications #Vector sum excited linear prediction #Voice activity detection #cs.SD