Byeongseon Park
- Phrase break prediction with bidirectional encoder representations in Japanese text-to-speech synthesis
2021/04/26 by Kosuke Futamata, Byeongseon Park, Futamata, Kosuke +5 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
- Audio-conditioned phonemic and prosodic annotation for building text-to-speech models from unlabeled speech data
2024/06/12 by Yuma Shirahata, Byeongseon Park, Shirahata, Yuma +5 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- Grapheme-Coherent Phonemic and Prosodic Annotation of Speech by Implicit and Explicit Grapheme Conditioning
2025/06/05 by Hien Ohnaka, Ohnaka, Hien, Yuma Shirahata +5 · 2 citations
Computer Science · Psychology · #Speech Recognition and Synthesis #Emotion and Mood Recognition #Phonetics and Phonology Research