Koyama, Yuichiro
- STARSS23: An Audio-Visual Dataset of Spatial Recordings of Real Scenes with Spatiotemporal Annotations of Sound Events
2023/06/15 by Kazuki Shimada, Shimada, Kazuki, Archontis Politis +21 · 25 citations
Biochemistry, Genetics and Molecular Biology · Computer Science · #Animal Vocal Communication and Behavior #Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Image and Video Processing (eess.IV) #Multimedia (cs.MM) #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- Multi-ACCDOA: Localizing and Detecting Overlapping Sounds from the Same Class with Auxiliary Duplicating Permutation Invariant Training
2021/10/14 by Shimada, Kazuki, Koyama, Yuichiro, Takahashi, Shusuke +3 · 16 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- ACCDOA: Activity-Coupled Cartesian Direction of Arrival Representation for Sound Event Localization and Detection
2020/10/29 by Kazuki Shimada, Yuichiro Koyama, Shimada, Kazuki +7 · 13 citations
Computer Science · #Music and Audio Processing #Speech and Audio Processing #Speech Recognition and Synthesis
- STARSS22: A dataset of spatial recordings of real scenes with spatiotemporal annotations of sound events
2022/06/04 by Archontis Politis, Kazuki Shimada, Politis, Archontis +17 · 11 citations
Computer Science · Health Professions · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Noise Effects and Management #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- Exploring the Best Loss Function for DNN-Based Low-latency Speech Enhancement with Temporal Convolutional Networks
2020/05/23 by Yuichiro Koyama, Koyama, Yuichiro, Tyler Vuong +5 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Zero- and Few-shot Sound Event Localization and Detection
2023/09/17 by Kazuki Shimada, Kengo Uchida, Shimada, Kazuki +11 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Distortion Audio Effects: Learning How to Recover the Clean Signal
2022/02/03 by Imort, Johannes, Fabbro, Giorgio, Ramírez, Marco A. Martínez +3 · 2 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- Ensemble of ACCDOA- and EINV2-based Systems with D3Nets and Impulse Response Simulation for Sound Event Localization and Detection
2021/06/21 by Kazuki Shimada, Naoya Takahashi, Shimada, Kazuki +11 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Diffusion-based Signal Refiner for Speech Separation
2023/05/10 by Hirano, Masato, Shimada, Kazuki, Koyama, Yuichiro +2 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Diffusion-Based Speech Enhancement with Joint Generative and Predictive Decoders
2023/05/18 by Shi, Hao, Shimada, Kazuki, Hirano, Masato +6 · 1 citation
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Music Foundation Model as Generic Booster for Music Downstream Tasks
2024/11/02 by Liao, WeiHsiang, Takida, Yuhta, Ikemiya, Yukara +13 · 3 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Information Retrieval (cs.IR) #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
- Spatial Data Augmentation with Simulated Room Impulse Responses for Sound Event Localization and Detection
2021/10/13 by Yuichiro Koyama, Kazuhide Shigemi, Koyama, Yuichiro +13 · 1 citation
Computer Science · Neuroscience · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Hearing Loss and Rehabilitation #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering