Sungwon Kim
- Glow-TTS: A Generative Flow for Text-to-Speech via Monotonic Alignment Search
2020/05/22 by Jaehyeon Kim, Sungwon Kim, Kim, Jaehyeon +5 · 29 citations
Computer Science · #Speech Recognition and Synthesis #Topic Modeling #Natural Language Processing Techniques
- Perception Prioritized Training of Diffusion Models
2022/04/01 by Jooyoung Choi, Jungbeom Lee, Choi, Jooyoung +9 · 25 citations
Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Music and Audio Processing
- Generative Adversarial Networks for Crystal Structure Prediction
2020/04/03 by Sungwon Kim, Kim, Sungwon, Juhwan Noh +7 · 8 citations
Materials Science · #Machine Learning in Materials Science #X-ray Diffraction in Crystallography #Electronic and Structural Properties of Oxides
- Interpretable Prototype-based Graph Information Bottleneck
2023/10/30 by Seo, Sangwoo, Sungwon Kim, Kim, Sungwon +2 · 4 citations
Computer Science · Materials Science · #Advanced Graph Neural Networks #Artificial Intelligence (cs.AI) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning in Materials Science
- Conditional Graph Information Bottleneck for Molecular Relational Learning
2023/04/29 by Namkyeong Lee, Lee, Namkyeong, Dongmin Hyun +9 · 2 citations
Computer Science · Materials Science · #Advanced Graph Neural Networks #Computational Drug Discovery Methods #FOS: Biological sciences #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning in Materials Science #Molecular Networks (q-bio.MN)
- VoiceTailor: Lightweight Plug-In Adapter for Diffusion-Based Personalized Text-to-Speech
2024/08/27 by Heeseung Kim, Sang-gil Lee, Kim, Heeseung +9 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- PersonaPlex: Voice and Role Control for Full Duplex Conversational Speech Models
2026/01/14 by Rajarshi Roy, Jonathan Raiman, Sang-gil Lee +5 · 1 voice · 1 citation
Computer Science · #cs.CL
- Audio-Visual Flamingo: Open Audio-Visual Intelligence for Long and Complex Videos
2026/07/17 by Sreyan Ghosh, Arushi Goel, Kaousheik Jayakumar +19
#eess.AS #cs.CV