vix.ing · top · new · best · stats · spec

Sungwon Kim

  1. Glow-TTS: A Generative Flow for Text-to-Speech via Monotonic Alignment Search
    2020/05/22 by Jaehyeon Kim, Sungwon Kim, Kim, Jaehyeon +5 · 29 citations
    Computer Science · #Speech Recognition and Synthesis #Topic Modeling #Natural Language Processing Techniques
  2. Perception Prioritized Training of Diffusion Models
    2022/04/01 by Jooyoung Choi, Jungbeom Lee, Choi, Jooyoung +9 · 25 citations
    Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Music and Audio Processing
  3. Generative Adversarial Networks for Crystal Structure Prediction
    2020/04/03 by Sungwon Kim, Kim, Sungwon, Juhwan Noh +7 · 8 citations
    Materials Science · #Machine Learning in Materials Science #X-ray Diffraction in Crystallography #Electronic and Structural Properties of Oxides
  4. Interpretable Prototype-based Graph Information Bottleneck
    2023/10/30 by Seo, Sangwoo, Sungwon Kim, Kim, Sungwon +2 · 4 citations
    Computer Science · Materials Science · #Advanced Graph Neural Networks #Artificial Intelligence (cs.AI) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning in Materials Science
  5. Conditional Graph Information Bottleneck for Molecular Relational Learning
    2023/04/29 by Namkyeong Lee, Lee, Namkyeong, Dongmin Hyun +9 · 2 citations
    Computer Science · Materials Science · #Advanced Graph Neural Networks #Computational Drug Discovery Methods #FOS: Biological sciences #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning in Materials Science #Molecular Networks (q-bio.MN)
  6. VoiceTailor: Lightweight Plug-In Adapter for Diffusion-Based Personalized Text-to-Speech
    2024/08/27 by Heeseung Kim, Sang-gil Lee, Kim, Heeseung +9 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
  7. PersonaPlex: Voice and Role Control for Full Duplex Conversational Speech Models
    2026/01/14 by Rajarshi Roy, Jonathan Raiman, Sang-gil Lee +5 · 1 voice · 1 citation
    Computer Science · #cs.CL
  8. Audio-Visual Flamingo: Open Audio-Visual Intelligence for Long and Complex Videos
    2026/07/17 by Sreyan Ghosh, Arushi Goel, Kaousheik Jayakumar +19
    #eess.AS #cs.CV