vix.ing · top · new · best · stats · spec

Gordon Wichern

  1. TF-Locoformer: Transformer with Local Modeling by Convolution for Speech Separation and Enhancement
    2024/08/06 by Kohei Saijo, Gordon Wichern, Saijo, Kohei +7 · 11 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  2. Cold Diffusion for Speech Enhancement
    2022/11/04 by Hao Yen, François G. Germain, Yen, Hao +5 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  3. Improving Audio Captioning Models with Fine-grained Audio Features, Text Embedding Supervision, and LLM Mix-up Augmentation
    2023/09/29 by Shih-Lun Wu, Wu, Shih-Lun, Xuankai Chang +11 · 3 citations
    Computer Science · Arts and Humanities · #Music and Audio Processing #Subtitles and Audiovisual Media #Speech Recognition and Synthesis
  4. Attentive Neural Processes and Batch Bayesian Optimization for Scalable\n Calibration of Physics-Informed Digital Twins
    2021/06/29 by Ankush Chakrabarty, Chakrabarty, Ankush, Gordon Wichern +3 · 2 citations
    Computer Science · Engineering · #Artificial Intelligence (cs.AI) #Building Energy and Comfort Optimization #Energy Load and Power Forecasting #FOS: Computer and information sciences #FOS: Mathematics #Gaussian Processes and Bayesian Inference #Industrial Vision Systems and Defect Detection #Machine Learning (cs.LG) #Neural Networks and Applications #Optimization and Control (math.OC)
  5. Why does music source separation benefit from cacophony?
    2024/02/28 by Chang-Bin Jeon, Gordon Wichern, Jeon, Chang-Bin +5 · 3 citations
    Neuroscience · Computer Science · #Hearing Loss and Rehabilitation #Music Technology and Sound Studies #Neuroscience and Music Perception
  6. The Cocktail Fork Problem: Three-Stem Audio Separation for Real-World Soundtracks
    2021/10/19 by Darius Petermann, Petermann, Darius, Gordon Wichern +5 · 2 citations
    Computer Science · #Speech and Audio Processing #Music and Audio Processing #Blind Source Separation Techniques
  7. NeuroHeed+: Improving Neuro-steered Speaker Extraction with Joint Auditory Attention Detection
    2023/12/12 by Zexu Pan, Gordon Wichern, Pan, Zexu +7 · 2 citations
    Computer Science · Neuroscience · #Audio and Speech Processing (eess.AS) #Blind Source Separation Techniques #EEG and Brain-Computer Interfaces #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  8. Task-Aware Unified Source Separation
    2024/10/31 by Kohei Saijo, Janek Ebbers, Saijo, Kohei +7 · 1 voice · 2 citations
    #eess.AS #cs.SD
  9. Retrieval-Augmented Neural Field for HRTF Upsampling and Personalization
    2025/01/22 by Yoshiki Masuyama, Masuyama, Yoshiki, Gordon Wichern +7 · 2 citations
    Business, Management and Accounting · #AI and HR Technologies #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  10. End-to-End Audio Visual Scene-Aware Dialog using Multimodal Attention-Based Video Features
    2018/06/21 by Chiori Hori, Hori, Chiori, Huda Alamri +23 · 1 citation
    Computer Science · #Multimodal Machine Learning Applications #Human Pose and Action Recognition #Video Analysis and Summarization
  11. Leveraging Audio-Only Data for Text-Queried Target Sound Extraction
    2024/09/20 by Kohei Saijo, Janek Ebbers, Saijo, Kohei +9 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  12. NABEATs: Noise-Aware Audio Representation Learning
    2026/07/18 by Takuya Fujimura, Yoshiki Masuyama, Gordon Wichern +3
    #eess.AS