vix.ing · top · new · best · stats · spec

Sato, Hiroshi

  1. How Bad Are Artifacts?: Analyzing the Impact of Speech Enhancement Errors on ASR
    2022/01/18 by Iwamoto, Kazuma, Ochiai, Tsubasa, Delcroix, Marc +4 · 5 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  2. Rethinking Processing Distortions: Disentangling the Impact of Speech Enhancement Errors on Speech Recognition Performance
    2024/04/23 by Ochiai, Tsubasa, Iwamoto, Kazuma, Delcroix, Marc +4 · 4 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  3. How does end-to-end speech recognition training impact speech enhancement artifacts?
    2023/11/20 by Iwamoto, Kazuma, Ochiai, Tsubasa, Delcroix, Marc +4 · 3 citations
    #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #electronic engineering #information engineering
  4. Smooth projective toric varieties whose nontrivial nef line bundles are big
    2008/10/23 by Osamu Fujino, Hiroshi Satō, Fujino, Osamu +1 · 1 citation
    Mathematics · #Algebraic Geometry and Number Theory #Advanced Algebra and Geometry #Algebraic structures and combinatorial models
  5. SpeakerBeam-SS: Real-time Target Speaker Extraction with Lightweight Conv-TasNet and State Space Modeling
    2024/07/01 by Hiroshi Sato, Sato, Hiroshi, Takafumi Moriya +15 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Signal Processing (eess.SP) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  6. End-to-End Joint Target and Non-Target Speakers ASR
    2023/06/04 by Ryo Masumura, Masumura, Ryo, Naoki Makishima +27 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
  7. Listen only to me! How well can target speech extraction handle false alarms?
    2022/04/11 by Marc Delcroix, Delcroix, Marc, Keisuke Kinoshita +9 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  8. Improving Scheduled Sampling for Neural Transducer-based ASR
    2023/05/25 by Moriya, Takafumi, Ashihara, Takanori, Sato, Hiroshi +3 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #electronic engineering #information engineering
  9. Streaming Target-Speaker ASR with Neural Transducer
    2022/09/09 by Takafumi Moriya, Moriya, Takafumi, Hiroshi Sato +7 · 1 citation
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
  10. Alignment-Free Training for Transducer-based Multi-Talker ASR
    2024/09/30 by Takafumi Moriya, Shota Horiguchi, Moriya, Takafumi +13 · 2 citations
    Engineering · Computer Science · #Ultrasonics and Acoustic Wave Propagation #Fault Detection and Control Systems #Speech Recognition and Synthesis
  11. NTT Multi-Speaker ASR System for the DASR Task of CHiME-8 Challenge
    2024/09/09 by Naoyuki Kamo, Naohiro Tawara, Kamo, Naoyuki +33 · 2 citations
    Computer Science · Engineering · #Advanced Data Processing Techniques #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Robotics and Automated Systems #Speech Recognition and Synthesis #electronic engineering #information engineering