vix.ing · top · new · best · stats · spec

Emiru Tsunoo

  1. Phoneme-aware Encoding for Prefix-tree-based Contextual ASR
    2023/12/15 by Hayato Futami, Emiru Tsunoo, Futami, Hayato +9 · 6 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  2. Transformer ASR with Contextual Block Processing
    2019/10/16 by Emiru Tsunoo, Tsunoo, Emiru, Yosuke Kashiwagi +5 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Speech Recognition and Synthesis #electronic engineering #information engineering
  3. UniverSLU: Universal Spoken Language Understanding for Diverse Tasks with Natural Language Instructions
    2023/10/04 by Siddhant Arora, Arora, Siddhant, Hayato Futami +13 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #Topic Modeling #electronic engineering #information engineering
  4. Ensemble of ACCDOA- and EINV2-based Systems with D3Nets and Impulse Response Simulation for Sound Event Localization and Detection
    2021/06/21 by Kazuki Shimada, Shimada, Kazuki, Naoya Takahashi +11 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  5. ESPnet-SDS: Unified Toolkit and Demo for Spoken Dialogue Systems
    2025/03/11 by Siddhant Arora, Yifan Peng, Arora, Siddhant +21 · 1 voice · 3 citations
    Computer Science · #Speech and dialogue systems #Multi-Agent Systems and Negotiation #Natural Language Processing Techniques
  6. Streaming Joint Speech Recognition and Disfluency Detection
    2022/11/16 by Hayato Futami, Emiru Tsunoo, Futami, Hayato +11 · 1 citation
    Computer Science · Psychology · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Phonetics and Phonology Research #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  7. Chain-of-Thought Training for Open E2E Spoken Dialogue Systems
    2025/05/31 by Siddhant Arora, Arora, Siddhant, Jinchuan Tian +13 · 5 citations
    Computer Science · #Speech and dialogue systems #Intelligent Tutoring Systems and Adaptive Learning #Topic Modeling
  8. Spatial Data Augmentation with Simulated Room Impulse Responses for Sound Event Localization and Detection
    2021/10/13 by Yuichiro Koyama, Kazuhide Shigemi, Koyama, Yuichiro +13 · 1 citation
    Computer Science · Neuroscience · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Hearing Loss and Rehabilitation #Music and Audio Processing #Sound (cs.SD) #Speech and Audio Processing #electronic engineering #information engineering
  9. Decoder-only Architecture for Streaming End-to-end Speech Recognition
    2024/06/23 by Emiru Tsunoo, Hayato Futami, Tsunoo, Emiru +7 · 2 citations
    Computer Science · #Advanced Data Compression Techniques #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  10. Causal Speech Enhancement with Predicting Semantics based on Quantized Self-supervised Learning Features
    2024/12/26 by Emiru Tsunoo, Tsunoo, Emiru, Yuki Saito +5 · 1 citation
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing
  11. Whale: Large-Scale multilingual ASR model with w2v-BERT and E-Branchformer with large speech data
    2025/06/02 by Yosuke Kashiwagi, Hayato Futami, Kashiwagi, Yosuke +5 · 2 citations
    Computer Science · #Speech Recognition and Synthesis