Shota Orihashi
- Enrollment-less training for personalized voice activity detection
2021/06/23 by Naoki Makishima, Makishima, Naoki, Mana Ihori +9 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Cross-Modal Transformer-Based Neural Correction Models for Automatic Speech Recognition
2021/07/04 by Tomohiro Tanaka, Tanaka, Tomohiro, Ryo Masumura +13 · 1 citation
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Music and Audio Processing #Natural Language Processing Techniques #Speech Recognition and Synthesis
- Unified Autoregressive Modeling for Joint End-to-End Multi-Talker Overlapped Speech Recognition and Speaker Attribute Estimation
2021/07/04 by Ryo Masumura, Masumura, Ryo, Daiki Okamura +11 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering