Emiru Tsunoo
- Phoneme-aware Encoding for Prefix-tree-based Contextual ASR
2023/12/15 by Hayato Futami, Emiru Tsunoo, Futami, Hayato +9 · 6 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
- Transformer ASR with Contextual Block Processing
2019/10/16 by Emiru Tsunoo, Tsunoo, Emiru, Yosuke Kashiwagi +5 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Speech Recognition and Synthesis #electronic engineering #information engineering
- UniverSLU: Universal Spoken Language Understanding for Diverse Tasks with Natural Language Instructions
2023/10/04 by Siddhant Arora, Arora, Siddhant, Hayato Futami +13 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #Topic Modeling #electronic engineering #information engineering
- Ensemble of ACCDOA- and EINV2-based Systems with D3Nets and Impulse Response Simulation for Sound Event Localization and Detection
2021/06/21 by Kazuki Shimada, Shimada, Kazuki, Naoya Takahashi +11 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- ESPnet-SDS: Unified Toolkit and Demo for Spoken Dialogue Systems
2025/03/11 by Siddhant Arora, Yifan Peng, Arora, Siddhant +21 · 1 voice · 3 citations
Computer Science · #Speech and dialogue systems #Multi-Agent Systems and Negotiation #Natural Language Processing Techniques
- Streaming Joint Speech Recognition and Disfluency Detection
2022/11/16 by Hayato Futami, Emiru Tsunoo, Futami, Hayato +11 · 1 citation
Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Chain-of-Thought Training for Open E2E Spoken Dialogue Systems
2025/05/31 by Siddhant Arora, Arora, Siddhant, Jinchuan Tian +13 · 5 citations
Computer Science · #Speech and dialogue systems #Intelligent Tutoring Systems and Adaptive Learning #Topic Modeling
- Spatial Data Augmentation with Simulated Room Impulse Responses for Sound Event Localization and Detection
2021/10/13 by Yuichiro Koyama, Kazuhide Shigemi, Koyama, Yuichiro +13 · 1 citation
Computer Science · Neuroscience · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Hearing Loss and Rehabilitation #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
- Decoder-only Architecture for Streaming End-to-end Speech Recognition
2024/06/23 by Emiru Tsunoo, Hayato Futami, Tsunoo, Emiru +7 · 2 citations
Computer Science · #Advanced Data Compression Techniques #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Causal Speech Enhancement with Predicting Semantics based on Quantized Self-supervised Learning Features
2024/12/26 by Emiru Tsunoo, Tsunoo, Emiru, Yuki Saito +5 · 1 citation
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing
- Whale: Large-Scale multilingual ASR model with w2v-BERT and E-Branchformer with large speech data
2025/06/02 by Yosuke Kashiwagi, Hayato Futami, Kashiwagi, Yosuke +5 · 2 citations
Computer Science · #Speech Recognition and Synthesis