Sato, Hiroshi
- How Bad Are Artifacts?: Analyzing the Impact of Speech Enhancement Errors on ASR
2022/01/18 by Iwamoto, Kazuma, Ochiai, Tsubasa, Delcroix, Marc +4 · 5 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Rethinking Processing Distortions: Disentangling the Impact of Speech Enhancement Errors on Speech Recognition Performance
2024/04/23 by Ochiai, Tsubasa, Iwamoto, Kazuma, Delcroix, Marc +4 · 4 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- How does end-to-end speech recognition training impact speech enhancement artifacts?
2023/11/20 by Iwamoto, Kazuma, Ochiai, Tsubasa, Delcroix, Marc +4 · 3 citations
#Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #electronic engineering #information engineering
- Smooth projective toric varieties whose nontrivial nef line bundles are big
2008/10/23 by Osamu Fujino, Hiroshi Satō, Fujino, Osamu +1 · 1 citation
Mathematics · #Algebraic Geometry and Number Theory #Advanced Algebra and Geometry #Algebraic structures and combinatorial models
- SpeakerBeam-SS: Real-time Target Speaker Extraction with Lightweight Conv-TasNet and State Space Modeling
2024/07/01 by Hiroshi Sato, Sato, Hiroshi, Takafumi Moriya +15 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Signal Processing (eess.SP) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- End-to-End Joint Target and Non-Target Speakers ASR
2023/06/04 by Ryo Masumura, Masumura, Ryo, Naoki Makishima +27 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
- Listen only to me! How well can target speech extraction handle false alarms?
2022/04/11 by Marc Delcroix, Delcroix, Marc, Keisuke Kinoshita +9 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Improving Scheduled Sampling for Neural Transducer-based ASR
2023/05/25 by Moriya, Takafumi, Ashihara, Takanori, Sato, Hiroshi +3 · 1 citation
#Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #electronic engineering #information engineering
- Streaming Target-Speaker ASR with Neural Transducer
2022/09/09 by Takafumi Moriya, Moriya, Takafumi, Hiroshi Sato +7 · 1 citation
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
- Alignment-Free Training for Transducer-based Multi-Talker ASR
2024/09/30 by Takafumi Moriya, Shota Horiguchi, Moriya, Takafumi +13 · 2 citations
Engineering · Computer Science · #Ultrasonics and Acoustic Wave Propagation #Fault Detection and Control Systems #Speech Recognition and Synthesis
- NTT Multi-Speaker ASR System for the DASR Task of CHiME-8 Challenge
2024/09/09 by Naoyuki Kamo, Naohiro Tawara, Kamo, Naoyuki +33 · 2 citations
Computer Science · Engineering · #Advanced Data Processing Techniques #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Robotics and Automated Systems #Speech Recognition and Synthesis #electronic engineering #information engineering