vix.ing · top · new · best · stats · spec

Byeongseon Park

  1. Phrase break prediction with bidirectional encoder representations in Japanese text-to-speech synthesis
    2021/04/26 by Kosuke Futamata, Byeongseon Park, Futamata, Kosuke +5 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  2. Audio-conditioned phonemic and prosodic annotation for building text-to-speech models from unlabeled speech data
    2024/06/12 by Yuma Shirahata, Byeongseon Park, Shirahata, Yuma +5 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
  3. Grapheme-Coherent Phonemic and Prosodic Annotation of Speech by Implicit and Explicit Grapheme Conditioning
    2025/06/05 by Hien Ohnaka, Ohnaka, Hien, Yuma Shirahata +5 · 2 citations
    Computer Science · Psychology · #Speech Recognition and Synthesis #Emotion and Mood Recognition #Phonetics and Phonology Research