vix.ing · top · new · best · stats · spec

Du, Xingjian

  1. HTS-AT: A Hierarchical Token-Semantic Audio Transformer for Sound Classification and Detection
    2022/02/02 by Ke Chen, Xingjian Du, Chen, Ke +9 · 27 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Information Retrieval (cs.IR) #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  2. Eagle and Finch: RWKV with Matrix-Valued States and Dynamic Recurrence
    2024/04/08 by Peng, Bo, Goldstein, Daniel, Anthony, Quentin +27 · 24 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
  3. RWKV-7 "Goose" with Expressive Dynamic State Evolution
    2025/03/18 by Peng, Bo, Zhang, Ruichong, Goldstein, Daniel +15 · 36 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #I.2.0 #I.2.7 #Machine Learning (cs.LG)
  4. Llasa: Scaling Train-Time and Inference-Time Compute for Llama-based Speech Synthesis
    2025/02/06 by Ye, Zhen, Zhu, Xinfa, Chan, Chi-Min +17 · 38 citations
    #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Multimedia (cs.MM) #Sound (cs.SD) #electronic engineering #information engineering
  5. YuE: Scaling Open Foundation Models for Long-Form Music Generation
    2025/03/11 by Ruibin Yuan, Yuan, Ruibin, Lin, Hanfeng +110 · 23 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computer Graphics and Visualization Techniques #FOS: Computer and information sciences #FOS: Electrical engineering #Multimedia (cs.MM) #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #electronic engineering #information engineering
  6. Universal Source Separation with Weakly Labelled Data
    2023/05/11 by Qiuqiang Kong, Kong, Qiuqiang, Ke Chen +11 · 6 citations
    Computer Science · #Speech and Audio Processing #Music and Audio Processing #Speech Recognition and Synthesis
  7. NotaGen: Advancing Musicality in Symbolic Music Generation with Large Language Model Training Paradigms
    2025/02/25 by Yashan Wang, Wang, Yashan, Shangda Wu +16 · 14 citations
    Computer Science · #Music and Audio Processing #Music Technology and Sound Studies
  8. Foundation Models for Music: A Survey
    2024/08/26 by Yinghao Ma, Ma, Yinghao, Anders Øland +81 · 8 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music Technology and Sound Studies #Sound (cs.SD) #electronic engineering #information engineering
  9. ByteCover3: Accurate Cover Song Identification on Short Queries
    2023/03/21 by Xingjian Du, Zijie J. Wang, Du, Xingjian +9 · 4 citations
    Computer Science · Arts and Humanities · #Music and Audio Processing #Diverse Musicological Studies #Music Technology and Sound Studies
  10. Zero-shot Audio Source Separation through Query-based Learning from Weakly-labeled Data
    2021/12/15 by Ke Chen, Xingjian Du, Chen, Ke +9 · 3 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Multimedia (cs.MM) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  11. ByteCover: Cover Song Identification via Multi-Loss Training
    2020/10/27 by Du, Xingjian, Yu, Zhesong, Zhu, Bilei +2 · 2 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  12. ByteComposer: a Human-like Melody Composition Method based on Language Model Agent
    2024/02/24 by Liang Xia, Liang, Xia, Du, Xingjian +5 · 3 citations
    Engineering · #Human Motion and Animation
  13. SymPAC: Scalable Symbolic Music Generation With Prompts And Constraints
    2024/09/04 by Haonan Chen, Jordan B. L. Smith, Chen, Haonan +12 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #FOS: Electrical engineering #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #electronic engineering #information engineering
  14. AudioTrust: Benchmarking the Multifaceted Trustworthiness of Audio Large Language Models
    2025/05/22 by Li, Kai, Shen, Can, Liu, Yile +31 · 5 citations
    #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  15. AAG-Stega: Automatic Audio Generation-based Steganography
    2018/09/10 by Zhongliang Yang, Xingjian Du, Yang, Zhongliang +7 · 1 citation
    Computer Science · #Advanced Steganography and Watermarking Techniques #Digital Media Forensic Detection #Music and Audio Processing
  16. Exploring Tokenization Methods for Multitrack Sheet Music Generation
    2024/10/23 by Yashan Wang, Wang, Yashan, Shangda Wu +5 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computer Graphics and Visualization Techniques #FOS: Computer and information sciences #FOS: Electrical engineering #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #electronic engineering #information engineering