vix.ing · top · new · best · stats · spec

Gustav Eje Henter

  1. Matcha-TTS: A fast TTS architecture with conditional flow matching
    2023/09/06 by Shivam Mehta, Ruibo Tu, Mehta, Shivam +7 · 71 citations
    Computer Science · #68T07 #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #H.5.5 #Human-Computer Interaction (cs.HC) #I.2.6 #I.2.7 #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  2. Listen, Denoise, Action! Audio-Driven Motion Synthesis with Diffusion Models
    2023/07/26 by Simon Alexanderson, Rajmund Nagy, Jonas Beskow +1 · 31 citations
    Engineering · Computer Science · #Human Motion and Animation #Human Pose and Action Recognition #Music and Audio Processing
  3. The GENEA Challenge 2023: A large scale evaluation of gesture generation models in monadic and dyadic settings
    2023/08/24 by Taras Kucherenko, Rajmund Nagy, Kucherenko, Taras +11 · 7 citations
    Computer Science · Psychology · #Speech and dialogue systems #Hand Gesture Recognition Systems #Social Robot Interaction and HRI
  4. MoGlow
    2019/05/31 by Gustav Eje Henter, Simon Alexanderson, Jonas Beskow · 3 citations
    Computer Science · Engineering · #Human Motion and Animation #Human Pose and Action Recognition #Video Analysis and Summarization
  5. The Case for Translation-Invariant Self-Attention in Transformer-Based\n Language Models
    2021/06/03 by Ulme Wennberg, Gustav Eje Henter, Wennberg, Ulme +1 · 3 citations
    Computer Science · #Topic Modeling #Natural Language Processing Techniques #Multimodal Machine Learning Applications
  6. Robust model training and generalisation with Studentising flows
    2020/06/11 by Simon Alexanderson, Alexanderson, Simon, Gustav Eje Henter +1 · 2 citations
    Computer Science · #62F35 (Secondary) #68T07 (Primary) #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #G.3 #Gaussian Processes and Bayesian Inference #Generative Adversarial Networks and Image Synthesis #I.2.6 #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  7. Should you use a probabilistic duration model in TTS? Probably! Especially for spontaneous speech
    2024/06/08 by Shivam Mehta, Harm Lameris, Mehta, Shivam +9 · 3 citations
    Computer Science · Psychology · #68T07 #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #H.5.5 #Human-Computer Interaction (cs.HC) #I.2.6 #I.2.7 #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  8. Voice Conversion-based Privacy through Adversarial Information Hiding
    2024/09/23 by Jacob J Webber, Webber, Jacob J, Oliver Watts +7 · 3 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #User Authentication and Security Systems
  9. Unified speech and gesture synthesis using flow matching
    2023/10/08 by Shivam Mehta, Ruibo Tu, Mehta, Shivam +9 · 2 citations
    Computer Science · #68T07 (Primary) #68T42 (Secondary) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Graphics (cs.GR) #H.5 #Hand Gesture Recognition Systems #Human-Computer Interaction (cs.HC) #I.2.6 #I.2.7 #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
  10. Deep Encoder-Decoder Models for Unsupervised Learning of Controllable Speech Synthesis
    2018/07/30 by Gustav Eje Henter, Henter, Gustav Eje, Jaime Lorenzo-Trueba +5 · 1 citation
    Computer Science · #62F99 #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #G.3 #I.2.7 #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  11. On the Use of Self-Supervised Speech Representations in Spontaneous Speech Synthesis
    2023/07/11 by Siyang Wang, Wang, Siyang, Gustav Eje Henter +5 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Human-Computer Interaction (cs.HC) #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  12. Causality for Tabular Data Synthesis: A High-Order Structure Causal Benchmark Framework
    2024/06/12 by Zineb Senane, Tu, Ruibo, Karlsson, Axel +12 · 1 citation
    Decision Sciences · #Scientific Computing and Data Management
  13. CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models
    2024/12/23 by Ruibo Tu, Hedvig Kjellström, Tu, Ruibo +5 · 1 citation
    Computer Science · #Topic Modeling #Natural Language Processing Techniques
  14. Fake it to make it: Using synthetic data to remedy the data shortage in joint multimodal speech-and-gesture synthesis
    2024/04/30 by Shivam Mehta, Mehta, Shivam, Anna Deichler +11 · 1 citation
    Computer Science · Psychology · #68T07 (Primary) #68T42 (Secondary) #Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Graphics (cs.GR) #H.5 #Hearing Impairment and Communication #Human-Computer Interaction (cs.HC) #I.2.6 #I.2.7 #Natural Language Processing Techniques #Sound (cs.SD) #Speech and dialogue systems #electronic engineering #information engineering