Higuchi, Yosuke
- Mask CTC: Non-Autoregressive End-to-End ASR with CTC and Mask Predict
2020/05/18 by Yosuke Higuchi, Higuchi, Yosuke, Shinji Watanabe +7 · 6 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Recent Developments on ESPnet Toolkit Boosted by Conformer
2020/10/26 by Guo, Pengcheng, Boyer, Florian, Chang, Xuankai +12 · 5 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Harnessing the Zero-Shot Power of Instruction-Tuned Large Language Model in End-to-End Speech Recognition
2023/09/19 by Yosuke Higuchi, Higuchi, Yosuke, Tetsuji Ogawa +3 · 1 voice · 1 citation
Computer Science · Engineering · #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #cs.CL #cs.SD #eess.AS
- Improving non-autoregressive end-to-end speech recognition with pre-trained acoustic and language models
2022/01/25 by Deng, Keqi, Yang, Zehui, Watanabe, Shinji +3 · 3 citations
#Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Orthros: Non-autoregressive End-to-end Speech Translation with Dual-decoder
2020/10/25 by Hirofumi Inaguma, Inaguma, Hirofumi, Yosuke Higuchi +7 · 4 citations
Computer Science · #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling
- Improved Mask-CTC for Non-Autoregressive End-to-End ASR
2020/10/26 by Higuchi, Yosuke, Inaguma, Hirofumi, Watanabe, Shinji +2 · 2 citations
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- A Comparative Study on Non-Autoregressive Modelings for Speech-to-Text Generation
2021/10/11 by Higuchi, Yosuke, Chen, Nanxin, Fujita, Yuya +6 · 2 citations
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Predictive Speech Recognition and End-of-Utterance Detection Towards Spoken Dialog Systems
2024/09/30 by Oswald Zink, Zink, Oswald, Yosuke Higuchi +7 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
- The 2020 ESPnet update: new features, broadened applications, performance improvements, and future plans
2020/12/23 by Shinji Watanabe, Florian Boyer, Watanabe, Shinji +27 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Non-autoregressive End-to-end Speech Translation with Parallel Autoregressive Rescoring
2021/09/09 by Hirofumi Inaguma, Inaguma, Hirofumi, Yosuke Higuchi +7 · 1 citation
Computer Science · #Natural Language Processing Techniques #Topic Modeling #Speech Recognition and Synthesis
- Hierarchical Conditional End-to-End ASR with CTC and Multi-Granular Subword Units
2021/10/08 by Higuchi, Yosuke, Karube, Keita, Ogawa, Tetsuji +1 · 1 citation
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #electronic engineering #information engineering
- CTC Alignments Improve Autoregressive Translation
2022/10/11 by Yan, Brian, Dalmia, Siddharth, Higuchi, Yosuke +4 · 1 citation
#Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- BECTRA: Transducer-based End-to-End ASR with BERT-Enhanced Encoder
2022/11/02 by Yosuke Higuchi, Higuchi, Yosuke, Tetsuji Ogawa +5 · 1 citation
Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Topic Modeling