vix.ing · top · new · best · stats · spec

Shutong Niu

  1. EmotiveTalk: Expressive Talking Head Generation through Audio Information Decoupling and Emotional Video Diffusion
    2024/11/23 by Haotian Wang, Wang, Haotian, Weng, Yuzhe +22 · 8 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Music and Audio Processing
  2. The USTC-NERCSLIP Systems for the CHiME-8 NOTSOFAR-1 Challenge
    2024/09/03 by Shutong Niu, Niu, Shutong, Ruoyu Wang +37 · 3 citations
    Earth and Planetary Sciences · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Geophysics and Gravity Measurements #Seismic Imaging and Inversion Techniques #Sound (cs.SD) #electronic engineering #information engineering
  3. The USTC-NERCSLIP Systems for the CHiME-7 DASR Challenge
    2023/08/28 by Ruoyu Wang, Maokui He, Wang, Ruoyu +35 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  4. The USTC-Ximalaya system for the ICASSP 2022 multi-channel multi-party meeting transcription (M2MeT) challenge
    2022/02/10 by Maokui He, He, Maokui, Xiang Lv +19 · 1 citation
    Computer Science · Social Sciences · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Public Relations and Crisis Communication #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  5. Semi-supervised multi-channel speaker diarization with cross-channel attention
    2023/07/17 by Shilong Wu, Wu, Shilong, Jun Du +11 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Speech Recognition and Synthesis #electronic engineering #information engineering
  6. Quality-Aware End-to-End Audio-Visual Neural Speaker Diarization
    2024/10/15 by Maokui He, He, Mao-Kui, Jun Du +7 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Multimedia (cs.MM) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  7. Neural Speaker Diarization Using Memory-Aware Multi-Speaker Embedding with Sequence-to-Sequence Architecture
    2023/09/17 by Gaobin Yang, Yang, Gaobin, Maokui He +15 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  8. Incorporating Spatial Cues in Modular Speaker Diarization for Multi-channel Multi-party Meetings
    2024/09/25 by Ruoyu Wang, Wang, Ruoyu, Shutong Niu +11 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering