vix.ing · top · new · best · stats · spec

Higuchi, Yosuke

  1. Mask CTC: Non-Autoregressive End-to-End ASR with CTC and Mask Predict
    2020/05/18 by Yosuke Higuchi, Higuchi, Yosuke, Shinji Watanabe +7 · 6 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  2. Recent Developments on ESPnet Toolkit Boosted by Conformer
    2020/10/26 by Guo, Pengcheng, Boyer, Florian, Chang, Xuankai +12 · 5 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  3. Harnessing the Zero-Shot Power of Instruction-Tuned Large Language Model in End-to-End Speech Recognition
    2023/09/19 by Yosuke Higuchi, Higuchi, Yosuke, Tetsuji Ogawa +3 · 1 voice · 1 citation
    Computer Science · Engineering · #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #cs.CL #cs.SD #eess.AS
  4. Improving non-autoregressive end-to-end speech recognition with pre-trained acoustic and language models
    2022/01/25 by Deng, Keqi, Yang, Zehui, Watanabe, Shinji +3 · 3 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  5. Orthros: Non-autoregressive End-to-end Speech Translation with Dual-decoder
    2020/10/25 by Hirofumi Inaguma, Inaguma, Hirofumi, Yosuke Higuchi +7 · 4 citations
    Computer Science · #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling
  6. Improved Mask-CTC for Non-Autoregressive End-to-End ASR
    2020/10/26 by Higuchi, Yosuke, Inaguma, Hirofumi, Watanabe, Shinji +2 · 2 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  7. A Comparative Study on Non-Autoregressive Modelings for Speech-to-Text Generation
    2021/10/11 by Higuchi, Yosuke, Chen, Nanxin, Fujita, Yuya +6 · 2 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  8. Predictive Speech Recognition and End-of-Utterance Detection Towards Spoken Dialog Systems
    2024/09/30 by Oswald Zink, Zink, Oswald, Yosuke Higuchi +7 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
  9. The 2020 ESPnet update: new features, broadened applications, performance improvements, and future plans
    2020/12/23 by Shinji Watanabe, Florian Boyer, Watanabe, Shinji +27 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  10. Non-autoregressive End-to-end Speech Translation with Parallel Autoregressive Rescoring
    2021/09/09 by Hirofumi Inaguma, Inaguma, Hirofumi, Yosuke Higuchi +7 · 1 citation
    Computer Science · #Natural Language Processing Techniques #Topic Modeling #Speech Recognition and Synthesis
  11. Hierarchical Conditional End-to-End ASR with CTC and Multi-Granular Subword Units
    2021/10/08 by Higuchi, Yosuke, Karube, Keita, Ogawa, Tetsuji +1 · 1 citation
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #electronic engineering #information engineering
  12. CTC Alignments Improve Autoregressive Translation
    2022/10/11 by Yan, Brian, Dalmia, Siddharth, Higuchi, Yosuke +4 · 1 citation
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  13. BECTRA: Transducer-based End-to-End ASR with BERT-Enhanced Encoder
    2022/11/02 by Yosuke Higuchi, Higuchi, Yosuke, Tetsuji Ogawa +5 · 1 citation
    Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Topic Modeling