Song, Eunwoo
- Parallel WaveGAN: A fast waveform generation model based on generative adversarial networks with multi-resolution spectrogram
2019/10/25 by Ryuichi Yamamoto, Eunwoo Song, Yamamoto, Ryuichi +3 · 63 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Signal Processing (eess.SP) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Paralinguistics-Aware Speech-Empowered Large Language Models for Natural Conversation
2024/02/08 by Heeseung Kim, Kim, Heeseung, Soonshin Seo +19 · 11 citations
Computer Science · #Speech and dialogue systems #Speech Recognition and Synthesis
- Improved parallel WaveGAN vocoder with perceptually weighted spectrogram loss
2021/01/19 by Eunwoo Song, Song, Eunwoo, Ryuichi Yamamoto +9 · 2 citations
Computer Science · #Advanced Data Compression Techniques #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Language Model-Based Emotion Prediction Methods for Emotional Speech Synthesis Systems
2022/06/30 by Hyun-Wook Yoon, Ohsung Kwon, Yoon, Hyun-Wook +11 · 2 citations
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Emotion and Mood Recognition #FOS: Computer and information sciences #FOS: Electrical engineering #Sentiment Analysis and Opinion Mining #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- TTS-by-TTS: TTS-driven Data Augmentation for Fast and High-Quality Speech Synthesis
2020/10/26 by Min-Jae Hwang, Hwang, Min-Jae, Ryuichi Yamamoto +5 · 2 citations
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
- Parallel waveform synthesis based on generative adversarial networks with voicing-aware conditional discriminators
2020/10/27 by Ryuichi Yamamoto, Eunwoo Song, Yamamoto, Ryuichi +5 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Signal Processing (eess.SP) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Cross-Speaker Emotion Transfer for Low-Resource Text-to-Speech Using Non-Parallel Voice Conversion with Pitch-Shift Data Augmentation
2022/04/21 by Ryo Terashima, Ryuichi Yamamoto, Terashima, Ryo +11 · 1 citation
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing
- Probability density distillation with generative adversarial networks for high-quality parallel waveform generation
2019/04/09 by Ryuichi Yamamoto, Yamamoto, Ryuichi, Eunwoo Song +3 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- LP-WaveNet: Linear Prediction-based WaveNet Speech Synthesis
2018/11/29 by Min-Jae Hwang, Frank K. Soong, Hwang, Min-Jae +9 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Period VITS: Variational Inference with Explicit Pitch Modeling for End-to-end Emotional Speech Synthesis
2022/10/28 by Yuma Shirahata, Ryuichi Yamamoto, Shirahata, Yuma +9 · 1 citation
Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Voice and Speech Disorders #electronic engineering #information engineering