vix.ing · top · new · best · stats · spec

King, Simon

  1. An Overview of Voice Conversion and its Challenges: From Statistical Modeling to Deep Learning
    2020/08/09 by Berrak Şişman, Junichi Yamagishi, Sisman, Berrak +5 · 29 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  2. Natural language guidance of high-fidelity text-to-speech with synthetic annotations
    2024/02/02 by Dan Lyth, Simon King, Lyth, Dan +1 · 34 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech and dialogue systems #electronic engineering #information engineering
  3. Investigating gated recurrent neural networks for speech synthesis
    2016/01/11 by Zhizheng Wu, Simon King, Wu, Zhizheng +1 · 1 voice
    Computer Science · #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #cs.CL #cs.NE
  4. Differentiable Grey-box Modelling of Phaser Effects using Frame-based Spectral Processing
    2023/06/02 by Alistair Carson, Carson, Alistair, Cassia Valentini-Botinhao +5 · 2 citations
    Computer Science · Engineering · #Image and Signal Denoising Methods #Structural Health Monitoring Techniques #Neural Networks and Applications
  5. The isomorphism problem for graded algebras and its application to mod-p cohomology rings of small p-groups
    2015/03/16 by Bettina Eick, Simon King, Eick, Bettina +1 · 1 citation
    Mathematics · #16Z05 (Primary) #20J06 (secondary) #Advanced Topics in Algebra #Algebraic structures and combinatorial models #Commutative Algebra and Its Applications #FOS: Mathematics #Rings and Algebras (math.RA) #math.RA #msc:16Z05 #msc:20J06
  6. Controllable Speaking Styles Using a Large Language Model
    2023/05/17 by Atli Thor Sigurgeirsson, Simon King, Sigurgeirsson, Atli Thor +1 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  7. Voice Conversion-based Privacy through Adversarial Information Hiding
    2024/09/23 by Jacob J Webber, Webber, Jacob J, Oliver Watts +7 · 3 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #User Authentication and Security Systems
  8. ADEPT: A Dataset for Evaluating Prosody Transfer
    2021/06/15 by Alexandra Torresquintero, Tian Huey Teh, Torresquintero, Alexandra +15 · 1 citation
    Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Phonetics and Phonology Research #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  9. Ctrl-P: Temporal Control of Prosodic Variation for Speech Synthesis
    2021/06/15 by Mohan, Devang S Ram, Hu, Vivian, Teh, Tian Huey +6 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  10. Do Prosody Transfer Models Transfer Prosody?
    2023/03/07 by Sigurgeirsson, Atli Thor, King, Simon · 1 citation
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  11. Do Discrete Self-Supervised Representations of Speech Capture Tone Distinctions?
    2024/10/25 by Opeyemi Osakuade, Simon King, Osakuade, Opeyemi +1 · 2 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing
  12. Can we reconstruct a dysarthric voice with the large speech model Parler TTS?
    2025/06/04 by Ariadna Sanchez, Simon King, Sanchez, Ariadna +1 · 1 voice · 1 citation
    Computer Science · Engineering · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #cs.CL #cs.SD #eess.AS #electronic engineering #information engineering