vix.ing · top · new · best · stats · spec

Qiquan Zhang

  1. Mamba in Speech: Towards an Alternative to Self-Attention
    2024/05/21 by Xiangyu Zhang, Zhang, Xiangyu, Qiquan Zhang +15 · 16 citations
    Social Sciences · #Audio and Speech Processing (eess.AS) #Education and Technology Integration #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  2. When LLMs Meets Acoustic Landmarks: An Efficient Approach to Integrate Speech into Large Language Models for Depression Detection
    2024/02/17 by Xiangyu Zhang, Zhang, Xiangyu, Hexin Liu +11 · 5 citations
    Psychology · Computer Science · #Mental Health via Writing #Topic Modeling
  3. FineD-Eval: Fine-grained Automatic Dialogue-Level Evaluation
    2022/10/25 by Chen Zhang, Luis Fernando D’Haro, Zhang, Chen +7 · 2 citations
    Computer Science · #Topic Modeling #Speech and dialogue systems #Multimodal Machine Learning Applications
  4. Speaking in Wavelet Domain: A Simple and Efficient Approach to Speed up Speech Diffusion Model
    2024/02/16 by Xiangyu Zhang, Daijiao Liu, Zhang, Xiangyu +13 · 3 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Speech and Audio Processing #electronic engineering #information engineering
  5. SAV-SE: Scene-aware Audio-Visual Speech Enhancement with Selective State Space Model
    2024/11/12 by Xinyuan Qian, Qian, Xinyuan, Yaodan Zhang +10 · 2 citations
    Computer Science · Engineering · #Advanced Adaptive Filtering Techniques #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Image and Signal Denoising Methods #Multimedia (cs.MM) #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  6. SpeechT-RAG: Reliable Depression Detection in LLMs with Retrieval-Augmented Generation Using Speech Timing Information
    2025/02/16 by Xiangyu Zhang, Zhang, Xiangyu, Hexin Liu +7 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  7. Selective State Space Model for Monaural Speech Enhancement
    2024/11/09 by Moran Chen, Qiquan Zhang, Chen, Moran +11 · 1 citation
    Computer Science · Engineering · #Advanced Adaptive Filtering Techniques #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering