vix.ing · top · new · best · stats · spec

Song, Eunwoo

  1. Parallel WaveGAN: A fast waveform generation model based on generative adversarial networks with multi-resolution spectrogram
    2019/10/25 by Ryuichi Yamamoto, Eunwoo Song, Yamamoto, Ryuichi +3 · 63 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Signal Processing (eess.SP) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  2. Paralinguistics-Aware Speech-Empowered Large Language Models for Natural Conversation
    2024/02/08 by Heeseung Kim, Kim, Heeseung, Soonshin Seo +19 · 11 citations
    Computer Science · #Speech and dialogue systems #Speech Recognition and Synthesis
  3. Improved parallel WaveGAN vocoder with perceptually weighted spectrogram loss
    2021/01/19 by Eunwoo Song, Song, Eunwoo, Ryuichi Yamamoto +9 · 2 citations
    Computer Science · #Advanced Data Compression Techniques #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  4. Language Model-Based Emotion Prediction Methods for Emotional Speech Synthesis Systems
    2022/06/30 by Hyun-Wook Yoon, Ohsung Kwon, Yoon, Hyun-Wook +11 · 2 citations
    Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Emotion and Mood Recognition #FOS: Computer and information sciences #FOS: Electrical engineering #Sentiment Analysis and Opinion Mining #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  5. TTS-by-TTS: TTS-driven Data Augmentation for Fast and High-Quality Speech Synthesis
    2020/10/26 by Min-Jae Hwang, Hwang, Min-Jae, Ryuichi Yamamoto +5 · 2 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
  6. Parallel waveform synthesis based on generative adversarial networks with voicing-aware conditional discriminators
    2020/10/27 by Ryuichi Yamamoto, Eunwoo Song, Yamamoto, Ryuichi +5 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Signal Processing (eess.SP) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  7. Cross-Speaker Emotion Transfer for Low-Resource Text-to-Speech Using Non-Parallel Voice Conversion with Pitch-Shift Data Augmentation
    2022/04/21 by Ryo Terashima, Ryuichi Yamamoto, Terashima, Ryo +11 · 1 citation
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing
  8. Probability density distillation with generative adversarial networks for high-quality parallel waveform generation
    2019/04/09 by Ryuichi Yamamoto, Yamamoto, Ryuichi, Eunwoo Song +3 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  9. LP-WaveNet: Linear Prediction-based WaveNet Speech Synthesis
    2018/11/29 by Min-Jae Hwang, Frank K. Soong, Hwang, Min-Jae +9 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  10. Period VITS: Variational Inference with Explicit Pitch Modeling for End-to-end Emotional Speech Synthesis
    2022/10/28 by Yuma Shirahata, Ryuichi Yamamoto, Shirahata, Yuma +9 · 1 citation
    Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Voice and Speech Disorders #electronic engineering #information engineering