Zheng‐Hua Tan
- An Overview of Deep-Learning-Based Audio-Visual Speech Enhancement and\n Separation
2020/08/21 by Daniel Michelsanti, Zheng‐Hua Tan, Michelsanti, Daniel +11 · 18 citations
Computer Science · Engineering · #Speech and Audio Processing #Advanced Adaptive Filtering Techniques #Blind Source Separation Techniques
- Multi-talker Speech Separation with Utterance-level Permutation\n Invariant Training of Deep Recurrent Neural Networks
2017/03/18 by Morten Kolbæk, Dong Yu, Kolbæk, Morten +5 · 12 citations
Computer Science · Psychology · #Speech and Audio Processing #Speech Recognition and Synthesis #Phonetics and Phonology Research
- Summary On The ICASSP 2022 Multi-Channel Multi-Party Meeting Transcription Grand Challenge
2022/02/08 by Fan Yu, Yu, Fan, Shiliang Zhang +29 · 4 citations
Computer Science · #Speech Recognition and Synthesis #Music and Audio Processing #Speech and Audio Processing
- Audio Mamba: Selective State Spaces for Self-Supervised Audio Representations
2024/06/04 by Sarthak Yadav, Zheng‐Hua Tan, Yadav, Sarthak +1 · 6 citations
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- Self-supervised Pretraining for Robust Personalized Voice Activity Detection in Adverse Conditions
2023/12/27 by Holger Severin Bovbjerg, Jesper Jensen, Bovbjerg, Holger Severin +5 · 3 citations
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
- Adversarial Multi-Task Deep Learning for Noise-Robust Voice Activity Detection with Low Algorithmic Delay
2022/07/04 by Claus Meyer Larsen, Larsen, Claus Meyer, Peter Koch +3 · 2 citations
Computer Science · #Anomaly Detection Techniques and Applications #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- BiSSL: Enhancing the Alignment Between Self-Supervised Pretraining and Downstream Fine-Tuning via Bilevel Optimization
2024/10/03 by Gustav Wagner Zakarias, Lars Kai Hansen, Zakarias, Gustav Wagner +3 · 3 citations
Engineering · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reservoir Engineering and Simulation Methods
- Hearing-Loss Compensation Using Deep Neural Networks: A Framework and Results From a Listening Test
2024/03/15 by Peter Leer, Jesper Jensen, Leer, Peter +9 · 2 citations
Health Professions · Neuroscience · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Hearing Loss and Rehabilitation #Noise Effects and Management #electronic engineering #information engineering
- Improvement of Noise-Robust Single-Channel Voice Activity Detection with Spatial Pre-processing
2021/04/12 by Max Væhrens, Væhrens, Max, Andreas Jonas Fuglsig +11 · 1 citation
Computer Science · #Speech and Audio Processing #Speech Recognition and Synthesis #Music and Audio Processing
- xLSTM-SENet: xLSTM for Single-Channel Speech Enhancement
2025/01/10 by Nikolai Lund Kühne, Kühne, Nikolai Lund, Jan Østergaard +6 · 2 voices · 4 citations
Computer Science · Health Professions · #Speech and Audio Processing #Speech Recognition and Synthesis #Infant Health and Development
- Adversarial Network Bottleneck Features for Noise Robust Speaker Verification
2017/06/11 by Hong Yu, Zheng‐Hua Tan, Yu, Hong +5 · 1 citation
Computer Science · #Speech and Audio Processing #Speech Recognition and Synthesis #Music and Audio Processing
- Improving Label-Deficient Keyword Spotting Through Self-Supervised Pretraining
2022/10/04 by Holger Severin Bovbjerg, Zheng‐Hua Tan, Bovbjerg, Holger Severin +1 · 1 citation
Computer Science · #68T10 #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Human-Computer Interaction (cs.HC) #I.2.6 #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- Noise-Robust Target-Speaker Voice Activity Detection Through Self-Supervised Pretraining
2025/01/06 by Holger Severin Bovbjerg, Jan Østergaard, Bovbjerg, Holger Severin +5 · 1 citation
Computer Science · #68T10 #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #I.2.6 #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Adversarial Example Detection by Classification for Deep Speech Recognition
2019/10/22 by Saeid Samizade, Zheng‐Hua Tan, Samizade, Saeid +5 · 1 citation
Computer Science · #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Network Security and Intrusion Detection #electronic engineering #information engineering