vix.ing · top · new · best · stats · spec

Skoglund, Jan

  1. SoundStream: An End-to-End Neural Audio Codec
    2021/07/07 by Zeghidour, Neil, Luebs, Alejandro, Omran, Ahmed +2 · 275 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  2. ViSQOL v3: An Open Source Production Ready Objective Speech and Audio Metric
    2020/04/20 by Michael Chinen, Felicia S. C. Lim, Chinen, Michael +9 · 46 citations
    Computer Science · Neuroscience · #Speech and Audio Processing #Hearing Loss and Rehabilitation #Image and Signal Denoising Methods
  3. LPCNet: Improving Neural Speech Synthesis Through Linear Prediction
    2018/10/28 by Valin, Jean-Marc, Skoglund, Jan · 20 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  4. SCOREQ: Speech Quality Assessment with Contrastive Regression
    2024/10/09 by Alessandro Ragano, Jan Skoglund, Ragano, Alessandro +3 · 1 voice · 24 citations
    Computer Science · Engineering · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #cs.SD #eess.AS #electronic engineering #information engineering
  5. Wavenet based low rate speech coding
    2017/12/01 by W. Bastiaan Kleijn, Felicia S. C. Lim, Kleijn, W. Bastiaan +11 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Signal Processing (eess.SP) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  6. NOMAD: Unsupervised Learning of Perceptual Embeddings for Speech Enhancement and Non-matching Reference Audio Quality Assessment
    2023/09/28 by Alessandro Ragano, Jan Skoglund, Ragano, Alessandro +3 · 6 citations
    Computer Science · Neuroscience · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Hearing Loss and Rehabilitation #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  7. Exploring Tradeoffs in Models for Low-latency Speech Enhancement
    2018/11/16 by Wilson, Kevin, Chinen, Michael, Thorpe, Jeremy +5 · 3 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  8. A Real-Time Wideband Neural Vocoder at 1.6 kb/s Using LPCNet
    2019/03/28 by Valin, Jean-Marc, Skoglund, Jan · 3 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  9. A Comparison of Deep Learning MOS Predictors for Speech Synthesis Quality
    2022/04/05 by Alessandro Ragano, Emmanouil Benetos, Ragano, Alessandro +11 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  10. WARP-Q: Quality Prediction For Generative Neural Speech Codecs
    2021/02/20 by Jassim, Wissam A., Skoglund, Jan, Chinen, Michael +1 · 2 citations
    #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Signal Processing (eess.SP) #electronic engineering #information engineering
  11. Generative Speech Coding with Predictive Variance Regularization
    2021/02/18 by Kleijn, W. Bastiaan, Storus, Andrew, Chinen, Michael +5 · 2 citations
    #94 #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #I.m #Sound (cs.SD) #electronic engineering #information engineering
  12. Ultra-Low-Bitrate Speech Coding with Pretrained Transformers
    2022/07/05 by Ali Siahkoohi, Siahkoohi, Ali, Michael Chinen +7 · 2 citations
    Computer Science · Engineering · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Geophysical Methods and Applications #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  13. LMCodec: A Low Bitrate Speech Codec With Causal Transformer Models
    2023/03/23 by Teerapat Jenrungrot, Jenrungrot, Teerapat, Michael Chinen +11 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  14. Improving Opus Low Bit Rate Quality with Neural Speech Synthesis
    2019/05/12 by Skoglund, Jan, Valin, Jean-Marc · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  15. BINAQUAL: A Full-Reference Objective Localization Similarity Metric for Binaural Audio
    2025/05/16 by Davoud Shariat Panah, Dan Barry, Panah, Davoud Shariat +7 · 4 citations
    Computer Science · Neuroscience · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Hearing Loss and Rehabilitation #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering