Yoshiki Masuyama
- Mamba-based Decoder-Only Approach with Bidirectional Speech Modeling for Speech Recognition
2024/11/11 by Yoshiki Masuyama, Masuyama, Yoshiki, Koichi Miyazaki +3 · 2 voices · 4 citations
#cs.SD #eess.AS
- End-to-End Integration of Speech Recognition, Dereverberation, Beamforming, and Self-Supervised Learning Representation
2022/10/19 by Yoshiki Masuyama, Xuankai Chang, Masuyama, Yoshiki +7 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Exploring the Capability of Mamba in Speech Applications
2024/06/24 by Koichi Miyazaki, Miyazaki, Koichi, Yoshiki Masuyama +3 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- ESPnet-SpeechLM: An Open Speech Language Model Toolkit
2025/02/21 by Jinchuan Tian, Tian, Jinchuan, Jiatong Shi +29 · 5 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech and dialogue systems #electronic engineering #information engineering
- Multi-Channel Target Speaker Extraction with Refinement: The WavLab Submission to the Second Clarity Enhancement Challenge
2023/02/15 by Samuele Cornell, Zhong-Qiu Wang, Cornell, Samuele +9 · 1 citation
Computer Science · Neuroscience · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Hearing Loss and Rehabilitation #Signal Processing (eess.SP) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Retrieval-Augmented Neural Field for HRTF Upsampling and Personalization
2025/01/22 by Yoshiki Masuyama, Gordon Wichern, Masuyama, Yoshiki +7 · 2 citations
Business, Management and Accounting · #AI and HR Technologies #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- NABEATs: Noise-Aware Audio Representation Learning
2026/07/18 by Takuya Fujimura, Yoshiki Masuyama, Gordon Wichern +3
#eess.AS