vix.ing · top · new · best · stats · spec

Zeng, Bang

  1. USEF-TSE: Universal Speaker Embedding Free Target Speaker Extraction
    2024/09/04 by Bang Zeng, Zeng, Bang, Ming Li +1 · 6 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  2. TSELM: Target Speaker Extraction using Discrete Tokens and Language Models
    2024/09/12 by Bo Tang, Bang Zeng, Tang, Beilong +3 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  3. LauraTSE: Target Speaker Extraction using Auto-Regressive Decoder-Only Language Models
    2025/04/10 by Tang, Beilong, Zeng, Bang, Li, Ming · 5 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  4. Universal Speaker Embedding Free Target Speaker Extraction and Personal Voice Activity Detection
    2025/01/07 by Zeng, Bang, Li, Ming · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering