Gustav Eje Henter
- Matcha-TTS: A fast TTS architecture with conditional flow matching
2023/09/06 by Shivam Mehta, Ruibo Tu, Mehta, Shivam +7 · 71 citations
Computer Science · #68T07 #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #H.5.5 #Human-Computer Interaction (cs.HC) #I.2.6 #I.2.7 #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Listen, Denoise, Action! Audio-Driven Motion Synthesis with Diffusion Models
2023/07/26 by Simon Alexanderson, Rajmund Nagy, Jonas Beskow +1 · 31 citations
Engineering · Computer Science · #Human Motion and Animation #Human Pose and Action Recognition #Music and Audio Processing
- The GENEA Challenge 2023: A large scale evaluation of gesture generation models in monadic and dyadic settings
2023/08/24 by Taras Kucherenko, Rajmund Nagy, Kucherenko, Taras +11 · 7 citations
Computer Science · Psychology · #Speech and dialogue systems #Hand Gesture Recognition Systems #Social Robot Interaction and HRI
- MoGlow
2019/05/31 by Gustav Eje Henter, Simon Alexanderson, Jonas Beskow · 3 citations
Computer Science · Engineering · #Human Motion and Animation #Human Pose and Action Recognition #Video Analysis and Summarization
- The Case for Translation-Invariant Self-Attention in Transformer-Based\n Language Models
2021/06/03 by Ulme Wennberg, Gustav Eje Henter, Wennberg, Ulme +1 · 3 citations
Computer Science · #Topic Modeling #Natural Language Processing Techniques #Multimodal Machine Learning Applications
- Robust model training and generalisation with Studentising flows
2020/06/11 by Simon Alexanderson, Alexanderson, Simon, Gustav Eje Henter +1 · 2 citations
Computer Science · #62F35 (Secondary) #68T07 (Primary) #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #G.3 #Gaussian Processes and Bayesian Inference #Generative Adversarial Networks and Image Synthesis #I.2.6 #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Should you use a probabilistic duration model in TTS? Probably! Especially for spontaneous speech
2024/06/08 by Shivam Mehta, Harm Lameris, Mehta, Shivam +9 · 3 citations
Computer Science · Psychology · #68T07 #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #H.5.5 #Human-Computer Interaction (cs.HC) #I.2.6 #I.2.7 #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Voice Conversion-based Privacy through Adversarial Information Hiding
2024/09/23 by Jacob J Webber, Webber, Jacob J, Oliver Watts +7 · 3 citations
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #User Authentication and Security Systems
- Unified speech and gesture synthesis using flow matching
2023/10/08 by Shivam Mehta, Ruibo Tu, Mehta, Shivam +9 · 2 citations
Computer Science · #68T07 (Primary) #68T42 (Secondary) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Graphics (cs.GR) #H.5 #Hand Gesture Recognition Systems #Human-Computer Interaction (cs.HC) #I.2.6 #I.2.7 #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- Deep Encoder-Decoder Models for Unsupervised Learning of Controllable Speech Synthesis
2018/07/30 by Gustav Eje Henter, Henter, Gustav Eje, Jaime Lorenzo-Trueba +5 · 1 citation
Computer Science · #62F99 #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #G.3 #I.2.7 #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- On the Use of Self-Supervised Speech Representations in Spontaneous Speech Synthesis
2023/07/11 by Siyang Wang, Wang, Siyang, Gustav Eje Henter +5 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Human-Computer Interaction (cs.HC) #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
- Causality for Tabular Data Synthesis: A High-Order Structure Causal Benchmark Framework
2024/06/12 by Zineb Senane, Tu, Ruibo, Karlsson, Axel +12 · 1 citation
Decision Sciences · #Scientific Computing and Data Management
- CARL-GT: Evaluating Causal Reasoning Capabilities of Large Language Models
2024/12/23 by Ruibo Tu, Hedvig Kjellström, Tu, Ruibo +5 · 1 citation
Computer Science · #Topic Modeling #Natural Language Processing Techniques
- Fake it to make it: Using synthetic data to remedy the data shortage in joint multimodal speech-and-gesture synthesis
2024/04/30 by Shivam Mehta, Mehta, Shivam, Anna Deichler +11 · 1 citation
Computer Science · Psychology · #68T07 (Primary) #68T42 (Secondary) #Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Graphics (cs.GR) #H.5 #Hearing Impairment and Communication #Human-Computer Interaction (cs.HC) #I.2.6 #I.2.7 #Natural Language Processing Techniques #Sound (cs.SD) #Speech and dialogue systems #electronic engineering #information engineering