vix.ing · top · new · best · stats · spec

Yoshiki Masuyama

  1. Mamba-based Decoder-Only Approach with Bidirectional Speech Modeling for Speech Recognition
    2024/11/11 by Yoshiki Masuyama, Masuyama, Yoshiki, Koichi Miyazaki +3 · 2 voices · 4 citations
    #cs.SD #eess.AS
  2. End-to-End Integration of Speech Recognition, Dereverberation, Beamforming, and Self-Supervised Learning Representation
    2022/10/19 by Yoshiki Masuyama, Xuankai Chang, Masuyama, Yoshiki +7 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  3. Exploring the Capability of Mamba in Speech Applications
    2024/06/24 by Koichi Miyazaki, Miyazaki, Koichi, Yoshiki Masuyama +3 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  4. ESPnet-SpeechLM: An Open Speech Language Model Toolkit
    2025/02/21 by Jinchuan Tian, Tian, Jinchuan, Jiatong Shi +29 · 5 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech and dialogue systems #electronic engineering #information engineering
  5. Multi-Channel Target Speaker Extraction with Refinement: The WavLab Submission to the Second Clarity Enhancement Challenge
    2023/02/15 by Samuele Cornell, Zhong-Qiu Wang, Cornell, Samuele +9 · 1 citation
    Computer Science · Neuroscience · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Hearing Loss and Rehabilitation #Signal Processing (eess.SP) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  6. Retrieval-Augmented Neural Field for HRTF Upsampling and Personalization
    2025/01/22 by Yoshiki Masuyama, Gordon Wichern, Masuyama, Yoshiki +7 · 2 citations
    Business, Management and Accounting · #AI and HR Technologies #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  7. NABEATs: Noise-Aware Audio Representation Learning
    2026/07/18 by Takuya Fujimura, Yoshiki Masuyama, Gordon Wichern +3
    #eess.AS