vix.ing · top · new · best · stats · spec

Jixun Yao

  1. GenSE: Generative Speech Enhancement via Language Models using Hierarchical Modeling
    2025/02/05 by Jixun Yao, Yao, Jixun, Hexin Liu +9 · 15 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
  2. Distinctive and Natural Speaker Anonymization via Singular Value Transformation-assisted Matrix
    2024/05/17 by Jixun Yao, Yao, Jixun, Qing Wang +7 · 7 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  3. PromptVC: Flexible Stylistic Voice Conversion in Latent Space Driven by Natural Language Prompts
    2023/09/17 by Jixun Yao, Yao, Jixun, Yuguang Yang +17 · 4 citations
    Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Voice and Speech Disorders #electronic engineering #information engineering
  4. StableVC: Style Controllable Zero-Shot Voice Conversion with Conditional Flow Matching
    2024/12/06 by Jixun Yao, Yuguang Yan, Yao, Jixun +11 · 7 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems
  5. MUSA: Multi-lingual Speaker Anonymization via Serial Disentanglement
    2024/07/16 by Jixun Yao, Yao, Jixun, Mengqing Wang +11 · 5 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  6. SongEval: A Benchmark Dataset for Song Aesthetics Evaluation
    2025/05/16 by Jixun Yao, Guobin Ma, Yao, Jixun +20 · 10 citations
    Computer Science · #Artificial Intelligence in Games #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Music Technology and Sound Studies #Music and Audio Processing #electronic engineering #information engineering
  7. UniSyn: An End-to-End Unified Model for Text-to-Speech and Singing Voice Synthesis
    2022/12/03 by Yi Lei, Lei, Yi, Shan Yang +11 · 2 citations
    Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Speech and Audio Processing
  8. DualVC 2: Dynamic Masked Convolution for Unified Streaming and Non-Streaming Voice Conversion
    2023/09/27 by Ziqian Ning, Yuepeng Jiang, Ning, Ziqian +11 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  9. StreamFlow: Streaming Flow Matching with Block-wise Guided Attention Mask for Speech Token Decoding
    2025/06/30 by Dake Guo, Guo, Dake, Jixun Yao +7 · 1 voice
    Computer Science · Engineering · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #cs.SD #eess.AS #electronic engineering #information engineering
  10. DiffRhythm+: Controllable and Flexible Full-Length Song Generation with Preference Optimization
    2025/07/17 by Yu-rou JIANG, Chen, Huakang, Jiang, Yuepeng +15 · 6 citations
    Computer Science · Engineering · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Human Motion and Animation #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #electronic engineering #information engineering
  11. SALT: Distinguishable Speaker Anonymization Through Latent Space Transformation
    2023/10/08 by Yuanjun Lv, Jixun Yao, Lv, Yuanjun +9 · 1 citation
    Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Voice and Speech Disorders #electronic engineering #information engineering
  12. Fine-grained Preference Optimization Improves Zero-shot Text-to-Speech
    2025/02/05 by Jixun Yao, Yao, Jixun, Yuguang Yang +13 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech and dialogue systems #electronic engineering #information engineering
  13. SongFormer: Scaling Music Structure Analysis with Heterogeneous Supervision
    2025/10/03 by C.X. Hao, Ruibin Yuan, Hao, Chunbo +10 · 2 citations
    Computer Science · #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing
  14. Towards Out-of-Distribution Detection in Vocoder Recognition via Latent Feature Reconstruction
    2024/06/04 by Renmingyue Du, Du, Renmingyue, Jixun Yao +5 · 1 citation
    #eess.AS
  15. Takin-VC: Expressive Zero-Shot Voice Conversion via Adaptive Hybrid Content Encoding and Enhanced Timbre Modeling
    2024/10/02 by Yuguang Yang, Yang, Yuguang, Yu Pan +15 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  16. Drop the beat! Freestyler for Accompaniment Conditioned Rapping Voice Generation
    2024/08/28 by Ziqian Ning, Shuai Wang, Ning, Ziqian +13 · 1 citation
    Computer Science · #Music Technology and Sound Studies #Speech and Audio Processing