vix.ing · top · new · best · stats · spec

Xihong Wu

  1. Cross-attention Inspired Selective State Space Models for Target Sound Extraction
    2024/09/07 by Donghang Wu, Wu, Donghang, Yiwen Wang +5 · 3 citations
    Computer Science · #Speech and Audio Processing #Speech Recognition and Synthesis #Music and Audio Processing
  2. TA-V2A: Textually Assisted Video-to-Audio Generation
    2025/03/12 by Yuhuan You, Xihong Wu, You, Yuhuan +3 · 2 citations
    Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Multimedia (cs.MM) #Multimodal Machine Learning Applications #Video Analysis and Summarization