vix.ing · top · new · best · stats · spec

Jinlong Xue

  1. Improving Audio Codec-based Zero-Shot Text-to-Speech Synthesis with Multi-Modal Context and Large Language Model
    2024/06/06 by Jinlong Xue, Yayue Deng, Xue, Jinlong +7 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  2. Frame-level emotional state alignment method for speech emotion recognition
    2023/12/27 by Qifei Li, Yingming Gao, Li, Qifei +11 · 1 citation
    Computer Science · Neuroscience · Psychology · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #EEG and Brain-Computer Interfaces #Emotion and Mood Recognition #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering