vix.ing · top · new · best · stats · spec

Tatsuya Kawahara

  1. Orthros: Non-autoregressive End-to-end Speech Translation with Dual-decoder
    2020/10/25 by Hirofumi Inaguma, Inaguma, Hirofumi, Yosuke Higuchi +7 · 4 citations
    Computer Science · #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling
  2. Real-time and Continuous Turn-taking Prediction Using Voice Activity Projection
    2024/01/10 by Koji Inoue, Bing’er Jiang, Inoue, Koji +7 · 4 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Human-Computer Interaction (cs.HC) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
  3. Zero- and Few-shot Sound Event Localization and Detection
    2023/09/17 by Kazuki Shimada, Kengo Uchida, Shimada, Kazuki +11 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  4. Speech Corpus of Ainu Folklore and End-to-end Speech Recognition for Ainu Language
    2020/02/16 by Kohei Matsuura, Matsuura, Kohei, Sei Ueno +7 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  5. Time-domain Speech Enhancement Assisted by Multi-resolution Frequency Encoder and Decoder
    2023/03/26 by Hao Shi, Masato Mimura, Shi, Hao +7 · 2 citations
    Computer Science · Engineering · #Advanced Adaptive Filtering Techniques #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  6. Multilingual End-to-End Speech Translation
    2019/10/01 by Hirofumi Inaguma, Inaguma, Hirofumi, Kevin Duh +5 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  7. Multilingual Turn-taking Prediction Using Voice Activity Projection
    2024/03/11 by Koji Inoue, Bing’er Jiang, Inoue, Koji +7 · 3 citations
    Computer Science · #Speech and dialogue systems
  8. Designing Precise and Robust Dialogue Response Evaluators
    2020/04/10 by Tianyu Zhao, Zhao, Tianyu, Divesh Lala +3 · 1 citation
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Speech and dialogue systems #Topic Modeling
  9. Alignment Knowledge Distillation for Online Streaming Attention-based Speech Recognition
    2021/02/28 by Hirofumi Inaguma, Tatsuya Kawahara, Inaguma, Hirofumi +1 · 1 citation
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
  10. Non-autoregressive End-to-end Speech Translation with Parallel Autoregressive Rescoring
    2021/09/09 by Hirofumi Inaguma, Yosuke Higuchi, Inaguma, Hirofumi +7 · 1 citation
    Computer Science · #Natural Language Processing Techniques #Topic Modeling #Speech Recognition and Synthesis
  11. Reasoning before Responding: Integrating Commonsense-based Causality Explanation for Empathetic Response Generation
    2023/07/28 by Yahui Fu, Koji Inoue, Fu, Yahui +5 · 2 citations
    Computer Science · Medicine · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Simulation-Based Education in Healthcare #Topic Modeling
  12. End-to-end Speech-to-Punctuated-Text Recognition
    2022/07/07 by Jumon Nozaki, Nozaki, Jumon, Tatsuya Kawahara +5 · 1 citation
    Computer Science · #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling
  13. Yeah, Un, Oh: Continuous and Real-time Backchannel Prediction with Fine-tuning of Voice Activity Projection
    2024/10/21 by Koji Inoue, Inoue, Koji, Divesh Lala +5 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Human-Computer Interaction (cs.HC) #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  14. A Noise-Robust Turn-Taking System for Real-World Dialogue Robots: A Field Experiment
    2025/03/08 by Koji Inoue, Yuki Okafuji, Inoue, Koji +9 · 1 citation
    Computer Science · Engineering · Psychology · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Robotics (cs.RO) #Robotics and Automated Systems #Social Robot Interaction and HRI #Sound (cs.SD) #Speech and dialogue systems
  15. Enhancing Long-term RAG Chatbots with Psychological Models of Memory Importance and Forgetting
    2024/09/19 by Ryuichi Sumida, Sumida, Ryuichi, Koji Inoue +3 · 1 citation
    Computer Science · Medicine · #AI in Service Interactions #Artificial Intelligence in Healthcare and Education
  16. Human-Like Embodied AI Interviewer: Employing Android ERICA in Real\n International Conference
    2024/12/13 by Zi Haur Pang, Yahui Fu, Pang, Zi Haur +9 · 1 citation
    Engineering · #Robotics and Automated Systems
  17. Serialized Speech Information Guidance with Overlapped Encoding Separation for Multi-Speaker Automatic Speech Recognition
    2024/09/01 by Hao Shi, Shi, Hao, Yuan Gao +5 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  18. Exploration of Adapter for Noise Robust Automatic Speech Recognition
    2024/02/28 by Hao Shi, Shi, Hao, Tatsuya Kawahara +1 · 1 citation
    Computer Science · #Speech Recognition and Synthesis
  19. Improving Zero-Shot Phonetic Classification through Language-Agnostic Articulatory Features
    2026/07/26 by Ryo Magoshi, Jaeyoung Lee, Shinsuke Sakai +1
    #cs.SD #eess.AS
  20. Memory-Driven Self-Disclosure and Relational Turning Points: A Longitudinal Multimodal Study of Human-AI Interaction
    2026/07/16 by Ryuichi Sumida, Mao Saeki, Masaki Eguchi +4
    Psychology · Computer Science · #Social Robot Interaction and HRI #Digital Mental Health Interventions #AI in Service Interactions
  21. On the Structure of Address in Multi-Party Dialogue: From Discrete Labels to Continuous Levels
    2026/07/17 by Taiga Mori, Koji Inoue, Divesh Lala +1
    #cs.CL #cs.AI #cs.HC