vix.ing · top · new · best · stats · spec

Wichern, Gordon

  1. WHAM!: Extending Speech Separation to Noisy Environments
    2019/07/02 by Wichern, Gordon, Antognini, Joe, Flynn, Michael +5 · 40 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Sound (cs.SD) #electronic engineering #information engineering
  2. WHAMR!: Noisy and Reverberant Single-Channel Speech Separation
    2019/10/22 by Maciejewski, Matthew, Wichern, Gordon, McQuinn, Emmett +1 · 14 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  3. Cutting Music Source Separation Some Slakh: A Dataset to Study the Impact of Training Data Quality and Quantity
    2019/09/18 by Manilow, Ethan, Wichern, Gordon, Seetharaman, Prem +1 · 9 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  4. STFT-Domain Neural Speech Enhancement with Very Low Algorithmic Latency
    2022/04/21 by Wang, Zhong-Qiu, Wichern, Gordon, Watanabe, Shinji +1 · 6 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  5. TF-Locoformer: Transformer with Local Modeling by Convolution for Speech Separation and Enhancement
    2024/08/06 by Kohei Saijo, Saijo, Kohei, Gordon Wichern +7 · 11 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  6. NIIRF: Neural IIR Filter Field for HRTF Upsampling and Personalization
    2024/02/27 by Masuyama, Yoshiki, Wichern, Gordon, Germain, François G. +4 · 5 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  7. Convolutive Prediction for Monaural Speech Dereverberation and Noisy-Reverberant Speaker Separation
    2021/08/16 by Wang, Zhong-Qiu, Wichern, Gordon, Roux, Jonathan Le · 4 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  8. Cold Diffusion for Speech Enhancement
    2022/11/04 by Hao Yen, Yen, Hao, François G. Germain +5 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  9. Class-conditional embeddings for music source separation
    2018/11/07 by Seetharaman, Prem, Wichern, Gordon, Venkataramani, Shrikant +1 · 2 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Sound (cs.SD) #electronic engineering #information engineering
  10. AutoClip: Adaptive Gradient Clipping for Source Separation Networks
    2020/07/25 by Seetharaman, Prem, Wichern, Gordon, Pardo, Bryan +1 · 2 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Sound (cs.SD) #electronic engineering #information engineering
  11. Improving Audio Captioning Models with Fine-grained Audio Features, Text Embedding Supervision, and LLM Mix-up Augmentation
    2023/09/29 by Shih-Lun Wu, Wu, Shih-Lun, Xuankai Chang +11 · 3 citations
    Computer Science · Arts and Humanities · #Music and Audio Processing #Subtitles and Audiovisual Media #Speech Recognition and Synthesis
  12. Attentive Neural Processes and Batch Bayesian Optimization for Scalable\n Calibration of Physics-Informed Digital Twins
    2021/06/29 by Ankush Chakrabarty, Chakrabarty, Ankush, Gordon Wichern +3 · 2 citations
    Computer Science · Engineering · #Artificial Intelligence (cs.AI) #Building Energy and Comfort Optimization #Energy Load and Power Forecasting #FOS: Computer and information sciences #FOS: Mathematics #Gaussian Processes and Bayesian Inference #Industrial Vision Systems and Defect Detection #Machine Learning (cs.LG) #Neural Networks and Applications #Optimization and Control (math.OC)
  13. Hyperbolic Audio Source Separation
    2022/12/09 by Petermann, Darius, Wichern, Gordon, Subramanian, Aswin +1 · 3 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  14. Why does music source separation benefit from cacophony?
    2024/02/28 by Chang-Bin Jeon, Jeon, Chang-Bin, Gordon Wichern +5 · 3 citations
    Neuroscience · Computer Science · #Hearing Loss and Rehabilitation #Music Technology and Sound Studies #Neuroscience and Music Perception
  15. 30+ Years of Source Separation Research: Achievements and Future Challenges
    2025/01/21 by Araki, Shoko, Ito, Nobutaka, Haeb-Umbach, Reinhold +3 · 5 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Signal Processing (eess.SP) #Sound (cs.SD) #electronic engineering #information engineering
  16. The Cocktail Fork Problem: Three-Stem Audio Separation for Real-World Soundtracks
    2021/10/19 by Darius Petermann, Gordon Wichern, Petermann, Darius +5 · 2 citations
    Computer Science · #Speech and Audio Processing #Music and Audio Processing #Blind Source Separation Techniques
  17. The Sound Demixing Challenge 2023 \unicodex2013 Cinematic Demixing Track
    2023/08/14 by Uhlich, Stefan, Fabbro, Giorgio, Hirano, Masato +14 · 2 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  18. Convolutive Prediction for Reverberant Speech Separation
    2021/08/16 by Wang, Zhong-Qiu, Wichern, Gordon, Roux, Jonathan Le · 2 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  19. Scenario-Aware Audio-Visual TF-GridNet for Target Speech Extraction
    2023/10/30 by Pan, Zexu, Wichern, Gordon, Masuyama, Yoshiki +4 · 2 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Multimedia (cs.MM) #electronic engineering #information engineering
  20. Bootstrapping single-channel source separation via unsupervised spatial clustering on stereo mixtures
    2018/11/06 by Seetharaman, Prem, Wichern, Gordon, Roux, Jonathan Le +1 · 1 citation
    #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Sound (cs.SD) #electronic engineering #information engineering
  21. NeuroHeed+: Improving Neuro-steered Speaker Extraction with Joint Auditory Attention Detection
    2023/12/12 by Zexu Pan, Gordon Wichern, Pan, Zexu +7 · 2 citations
    Computer Science · Neuroscience · #Audio and Speech Processing (eess.AS) #Blind Source Separation Techniques #EEG and Brain-Computer Interfaces #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  22. Generation or Replication: Auscultating Audio Latent Diffusion Models
    2023/10/16 by Bralios, Dimitrios, Wichern, Gordon, Germain, François G. +4 · 2 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  23. Meta-Learning of Neural State-Space Models Using Data From Similar Systems
    2022/11/14 by Chakrabarty, Ankush, Wichern, Gordon, Laughman, Christopher R. · 1 citation
    #FOS: Computer and information sciences #FOS: Electrical engineering #FOS: Mathematics #Machine Learning (cs.LG) #Optimization and Control (math.OC) #Systems and Control (eess.SY) #electronic engineering #information engineering
  24. Late Audio-Visual Fusion for In-The-Wild Speaker Diarization
    2022/11/02 by Pan, Zexu, Wichern, Gordon, Germain, François G. +2 · 1 citation
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  25. Task-Aware Unified Source Separation
    2024/10/31 by Kohei Saijo, Saijo, Kohei, Janek Ebbers +7 · 1 voice · 2 citations
    #eess.AS #cs.SD
  26. Leveraging Low-Distortion Target Estimates for Improved Speech Enhancement
    2021/10/01 by Wang, Zhong-Qiu, Wichern, Gordon, Roux, Jonathan Le · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  27. Tackling the Cocktail Fork Problem for Separation and Transcription of Real-World Soundtracks
    2022/12/14 by Petermann, Darius, Wichern, Gordon, Subramanian, Aswin Shanmugam +2 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  28. Retrieval-Augmented Neural Field for HRTF Upsampling and Personalization
    2025/01/22 by Yoshiki Masuyama, Masuyama, Yoshiki, Gordon Wichern +7 · 2 citations
    Business, Management and Accounting · #AI and HR Technologies #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  29. Sound Event Bounding Boxes
    2024/06/06 by Ebbers, Janek, Germain, Francois G., Wichern, Gordon +1 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  30. End-to-End Audio Visual Scene-Aware Dialog using Multimodal Attention-Based Video Features
    2018/06/21 by Chiori Hori, Hori, Chiori, Huda Alamri +23 · 1 citation
    Computer Science · #Multimodal Machine Learning Applications #Human Pose and Action Recognition #Video Analysis and Summarization
  31. Leveraging Audio-Only Data for Text-Queried Target Sound Extraction
    2024/09/20 by Kohei Saijo, Janek Ebbers, Saijo, Kohei +9 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  32. Data Augmentation Using Neural Acoustic Fields With Retrieval-Augmented Pre-training
    2025/04/19 by Ick, Christopher, Wichern, Gordon, Masuyama, Yoshiki +2 · 1 citation
    #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering