Wichern, Gordon
- WHAM!: Extending Speech Separation to Noisy Environments
2019/07/02 by Wichern, Gordon, Antognini, Joe, Flynn, Michael +5 · 40 citations
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Sound (cs.SD) #electronic engineering #information engineering
- WHAMR!: Noisy and Reverberant Single-Channel Speech Separation
2019/10/22 by Maciejewski, Matthew, Wichern, Gordon, McQuinn, Emmett +1 · 14 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Cutting Music Source Separation Some Slakh: A Dataset to Study the Impact of Training Data Quality and Quantity
2019/09/18 by Manilow, Ethan, Wichern, Gordon, Seetharaman, Prem +1 · 9 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- STFT-Domain Neural Speech Enhancement with Very Low Algorithmic Latency
2022/04/21 by Wang, Zhong-Qiu, Wichern, Gordon, Watanabe, Shinji +1 · 6 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- TF-Locoformer: Transformer with Local Modeling by Convolution for Speech Separation and Enhancement
2024/08/06 by Kohei Saijo, Saijo, Kohei, Gordon Wichern +7 · 11 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- NIIRF: Neural IIR Filter Field for HRTF Upsampling and Personalization
2024/02/27 by Masuyama, Yoshiki, Wichern, Gordon, Germain, François G. +4 · 5 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Convolutive Prediction for Monaural Speech Dereverberation and Noisy-Reverberant Speaker Separation
2021/08/16 by Wang, Zhong-Qiu, Wichern, Gordon, Roux, Jonathan Le · 4 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Cold Diffusion for Speech Enhancement
2022/11/04 by Hao Yen, Yen, Hao, François G. Germain +5 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Class-conditional embeddings for music source separation
2018/11/07 by Seetharaman, Prem, Wichern, Gordon, Venkataramani, Shrikant +1 · 2 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Sound (cs.SD) #electronic engineering #information engineering
- AutoClip: Adaptive Gradient Clipping for Source Separation Networks
2020/07/25 by Seetharaman, Prem, Wichern, Gordon, Pardo, Bryan +1 · 2 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Sound (cs.SD) #electronic engineering #information engineering
- Improving Audio Captioning Models with Fine-grained Audio Features, Text Embedding Supervision, and LLM Mix-up Augmentation
2023/09/29 by Shih-Lun Wu, Wu, Shih-Lun, Xuankai Chang +11 · 3 citations
Computer Science · Arts and Humanities · #Music and Audio Processing #Subtitles and Audiovisual Media #Speech Recognition and Synthesis
- Attentive Neural Processes and Batch Bayesian Optimization for Scalable\n Calibration of Physics-Informed Digital Twins
2021/06/29 by Ankush Chakrabarty, Chakrabarty, Ankush, Gordon Wichern +3 · 2 citations
Computer Science · Engineering · #Artificial Intelligence (cs.AI) #Building Energy and Comfort Optimization #Energy Load and Power Forecasting #FOS: Computer and information sciences #FOS: Mathematics #Gaussian Processes and Bayesian Inference #Industrial Vision Systems and Defect Detection #Machine Learning (cs.LG) #Neural Networks and Applications #Optimization and Control (math.OC)
- Hyperbolic Audio Source Separation
2022/12/09 by Petermann, Darius, Wichern, Gordon, Subramanian, Aswin +1 · 3 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Why does music source separation benefit from cacophony?
2024/02/28 by Chang-Bin Jeon, Jeon, Chang-Bin, Gordon Wichern +5 · 3 citations
Neuroscience · Computer Science · #Hearing Loss and Rehabilitation #Music Technology and Sound Studies #Neuroscience and Music Perception
- 30+ Years of Source Separation Research: Achievements and Future Challenges
2025/01/21 by Araki, Shoko, Ito, Nobutaka, Haeb-Umbach, Reinhold +3 · 5 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Signal Processing (eess.SP) #Sound (cs.SD) #electronic engineering #information engineering
- The Cocktail Fork Problem: Three-Stem Audio Separation for Real-World Soundtracks
2021/10/19 by Darius Petermann, Gordon Wichern, Petermann, Darius +5 · 2 citations
Computer Science · #Speech and Audio Processing #Music and Audio Processing #Blind Source Separation Techniques
- The Sound Demixing Challenge 2023 \unicodex2013 Cinematic Demixing Track
2023/08/14 by Uhlich, Stefan, Fabbro, Giorgio, Hirano, Masato +14 · 2 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Convolutive Prediction for Reverberant Speech Separation
2021/08/16 by Wang, Zhong-Qiu, Wichern, Gordon, Roux, Jonathan Le · 2 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Scenario-Aware Audio-Visual TF-GridNet for Target Speech Extraction
2023/10/30 by Pan, Zexu, Wichern, Gordon, Masuyama, Yoshiki +4 · 2 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Multimedia (cs.MM) #electronic engineering #information engineering
- Bootstrapping single-channel source separation via unsupervised spatial clustering on stereo mixtures
2018/11/06 by Seetharaman, Prem, Wichern, Gordon, Roux, Jonathan Le +1 · 1 citation
#Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Sound (cs.SD) #electronic engineering #information engineering
- NeuroHeed+: Improving Neuro-steered Speaker Extraction with Joint Auditory Attention Detection
2023/12/12 by Zexu Pan, Gordon Wichern, Pan, Zexu +7 · 2 citations
Computer Science · Neuroscience · #Audio and Speech Processing (eess.AS) #Blind Source Separation Techniques #EEG and Brain-Computer Interfaces #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- Generation or Replication: Auscultating Audio Latent Diffusion Models
2023/10/16 by Bralios, Dimitrios, Wichern, Gordon, Germain, François G. +4 · 2 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Meta-Learning of Neural State-Space Models Using Data From Similar Systems
2022/11/14 by Chakrabarty, Ankush, Wichern, Gordon, Laughman, Christopher R. · 1 citation
#FOS: Computer and information sciences #FOS: Electrical engineering #FOS: Mathematics #Machine Learning (cs.LG) #Optimization and Control (math.OC) #Systems and Control (eess.SY) #electronic engineering #information engineering
- Late Audio-Visual Fusion for In-The-Wild Speaker Diarization
2022/11/02 by Pan, Zexu, Wichern, Gordon, Germain, François G. +2 · 1 citation
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Task-Aware Unified Source Separation
2024/10/31 by Kohei Saijo, Saijo, Kohei, Janek Ebbers +7 · 1 voice · 2 citations
#eess.AS #cs.SD
- Leveraging Low-Distortion Target Estimates for Improved Speech Enhancement
2021/10/01 by Wang, Zhong-Qiu, Wichern, Gordon, Roux, Jonathan Le · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Tackling the Cocktail Fork Problem for Separation and Transcription of Real-World Soundtracks
2022/12/14 by Petermann, Darius, Wichern, Gordon, Subramanian, Aswin Shanmugam +2 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Retrieval-Augmented Neural Field for HRTF Upsampling and Personalization
2025/01/22 by Yoshiki Masuyama, Masuyama, Yoshiki, Gordon Wichern +7 · 2 citations
Business, Management and Accounting · #AI and HR Technologies #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Sound Event Bounding Boxes
2024/06/06 by Ebbers, Janek, Germain, Francois G., Wichern, Gordon +1 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- End-to-End Audio Visual Scene-Aware Dialog using Multimodal Attention-Based Video Features
2018/06/21 by Chiori Hori, Hori, Chiori, Huda Alamri +23 · 1 citation
Computer Science · #Multimodal Machine Learning Applications #Human Pose and Action Recognition #Video Analysis and Summarization
- Leveraging Audio-Only Data for Text-Queried Target Sound Extraction
2024/09/20 by Kohei Saijo, Janek Ebbers, Saijo, Kohei +9 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Data Augmentation Using Neural Acoustic Fields With Retrieval-Augmented Pre-training
2025/04/19 by Ick, Christopher, Wichern, Gordon, Masuyama, Yoshiki +2 · 1 citation
#Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering