vix.ing · top · new · best · stats · spec

Zheng‐Hua Tan

  1. An Overview of Deep-Learning-Based Audio-Visual Speech Enhancement and\n Separation
    2020/08/21 by Daniel Michelsanti, Zheng‐Hua Tan, Michelsanti, Daniel +11 · 18 citations
    Computer Science · Engineering · #Speech and Audio Processing #Advanced Adaptive Filtering Techniques #Blind Source Separation Techniques
  2. Multi-talker Speech Separation with Utterance-level Permutation\n Invariant Training of Deep Recurrent Neural Networks
    2017/03/18 by Morten Kolbæk, Dong Yu, Kolbæk, Morten +5 · 12 citations
    Computer Science · Psychology · #Speech and Audio Processing #Speech Recognition and Synthesis #Phonetics and Phonology Research
  3. Summary On The ICASSP 2022 Multi-Channel Multi-Party Meeting Transcription Grand Challenge
    2022/02/08 by Fan Yu, Yu, Fan, Shiliang Zhang +29 · 4 citations
    Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Speech and Audio Processing
  4. Audio Mamba: Selective State Spaces for Self-Supervised Audio Representations
    2024/06/04 by Sarthak Yadav, Zheng‐Hua Tan, Yadav, Sarthak +1 · 6 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  5. Self-supervised Pretraining for Robust Personalized Voice Activity Detection in Adverse Conditions
    2023/12/27 by Holger Severin Bovbjerg, Jesper Jensen, Bovbjerg, Holger Severin +5 · 3 citations
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
  6. Adversarial Multi-Task Deep Learning for Noise-Robust Voice Activity Detection with Low Algorithmic Delay
    2022/07/04 by Claus Meyer Larsen, Larsen, Claus Meyer, Peter Koch +3 · 2 citations
    Computer Science · #Anomaly Detection Techniques and Applications #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  7. BiSSL: Enhancing the Alignment Between Self-Supervised Pretraining and Downstream Fine-Tuning via Bilevel Optimization
    2024/10/03 by Gustav Wagner Zakarias, Lars Kai Hansen, Zakarias, Gustav Wagner +3 · 3 citations
    Engineering · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reservoir Engineering and Simulation Methods
  8. Hearing-Loss Compensation Using Deep Neural Networks: A Framework and Results From a Listening Test
    2024/03/15 by Peter Leer, Jesper Jensen, Leer, Peter +9 · 2 citations
    Health Professions · Neuroscience · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Hearing Loss and Rehabilitation #Noise Effects and Management #electronic engineering #information engineering
  9. Improvement of Noise-Robust Single-Channel Voice Activity Detection with Spatial Pre-processing
    2021/04/12 by Max Væhrens, Væhrens, Max, Andreas Jonas Fuglsig +11 · 1 citation
    Computer Science · #Speech and Audio Processing #Speech Recognition and Synthesis #Music and Audio Processing
  10. xLSTM-SENet: xLSTM for Single-Channel Speech Enhancement
    2025/01/10 by Nikolai Lund Kühne, Kühne, Nikolai Lund, Jan Østergaard +6 · 2 voices · 4 citations
    Computer Science · Health Professions · #Speech and Audio Processing #Speech Recognition and Synthesis #Infant Health and Development
  11. Adversarial Network Bottleneck Features for Noise Robust Speaker Verification
    2017/06/11 by Hong Yu, Zheng‐Hua Tan, Yu, Hong +5 · 1 citation
    Computer Science · #Speech and Audio Processing #Speech Recognition and Synthesis #Music and Audio Processing
  12. Improving Label-Deficient Keyword Spotting Through Self-Supervised Pretraining
    2022/10/04 by Holger Severin Bovbjerg, Zheng‐Hua Tan, Bovbjerg, Holger Severin +1 · 1 citation
    Computer Science · #68T10 #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Human-Computer Interaction (cs.HC) #I.2.6 #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
  13. Noise-Robust Target-Speaker Voice Activity Detection Through Self-Supervised Pretraining
    2025/01/06 by Holger Severin Bovbjerg, Jan Østergaard, Bovbjerg, Holger Severin +5 · 1 citation
    Computer Science · #68T10 #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #I.2.6 #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  14. Adversarial Example Detection by Classification for Deep Speech Recognition
    2019/10/22 by Saeid Samizade, Zheng‐Hua Tan, Samizade, Saeid +5 · 1 citation
    Computer Science · #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Network Security and Intrusion Detection #electronic engineering #information engineering