vix.ing · top · new · best · stats · spec

Koyama, Yuichiro

  1. STARSS23: An Audio-Visual Dataset of Spatial Recordings of Real Scenes with Spatiotemporal Annotations of Sound Events
    2023/06/15 by Kazuki Shimada, Shimada, Kazuki, Archontis Politis +21 · 25 citations
    Biochemistry, Genetics and Molecular Biology · Computer Science · #Animal Vocal Communication and Behavior #Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Image and Video Processing (eess.IV) #Multimedia (cs.MM) #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  2. Multi-ACCDOA: Localizing and Detecting Overlapping Sounds from the Same Class with Auxiliary Duplicating Permutation Invariant Training
    2021/10/14 by Shimada, Kazuki, Koyama, Yuichiro, Takahashi, Shusuke +3 · 16 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  3. ACCDOA: Activity-Coupled Cartesian Direction of Arrival Representation for Sound Event Localization and Detection
    2020/10/29 by Kazuki Shimada, Yuichiro Koyama, Shimada, Kazuki +7 · 13 citations
    Computer Science · #Music and Audio Processing #Speech and Audio Processing #Speech Recognition and Synthesis
  4. STARSS22: A dataset of spatial recordings of real scenes with spatiotemporal annotations of sound events
    2022/06/04 by Archontis Politis, Kazuki Shimada, Politis, Archontis +17 · 11 citations
    Computer Science · Health Professions · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Noise Effects and Management #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  5. Exploring the Best Loss Function for DNN-Based Low-latency Speech Enhancement with Temporal Convolutional Networks
    2020/05/23 by Yuichiro Koyama, Koyama, Yuichiro, Tyler Vuong +5 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  6. Zero- and Few-shot Sound Event Localization and Detection
    2023/09/17 by Kazuki Shimada, Kengo Uchida, Shimada, Kazuki +11 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  7. Distortion Audio Effects: Learning How to Recover the Clean Signal
    2022/02/03 by Imort, Johannes, Fabbro, Giorgio, Ramírez, Marco A. Martínez +3 · 2 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  8. Ensemble of ACCDOA- and EINV2-based Systems with D3Nets and Impulse Response Simulation for Sound Event Localization and Detection
    2021/06/21 by Kazuki Shimada, Naoya Takahashi, Shimada, Kazuki +11 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  9. Diffusion-based Signal Refiner for Speech Separation
    2023/05/10 by Hirano, Masato, Shimada, Kazuki, Koyama, Yuichiro +2 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  10. Diffusion-Based Speech Enhancement with Joint Generative and Predictive Decoders
    2023/05/18 by Shi, Hao, Shimada, Kazuki, Hirano, Masato +6 · 1 citation
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  11. Music Foundation Model as Generic Booster for Music Downstream Tasks
    2024/11/02 by Liao, WeiHsiang, Takida, Yuhta, Ikemiya, Yukara +13 · 3 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Information Retrieval (cs.IR) #Machine Learning (cs.LG) #Sound (cs.SD) #electronic engineering #information engineering
  12. Spatial Data Augmentation with Simulated Room Impulse Responses for Sound Event Localization and Detection
    2021/10/13 by Yuichiro Koyama, Kazuhide Shigemi, Koyama, Yuichiro +13 · 1 citation
    Computer Science · Neuroscience · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Hearing Loss and Rehabilitation #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering