Tatsuya Kawahara
- Orthros: Non-autoregressive End-to-end Speech Translation with Dual-decoder
2020/10/25 by Hirofumi Inaguma, Inaguma, Hirofumi, Yosuke Higuchi +7 · 4 citations
Computer Science · #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling
- Real-time and Continuous Turn-taking Prediction Using Voice Activity Projection
2024/01/10 by Koji Inoue, Bing’er Jiang, Inoue, Koji +7 · 4 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Human-Computer Interaction (cs.HC) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
- Zero- and Few-shot Sound Event Localization and Detection
2023/09/17 by Kazuki Shimada, Kengo Uchida, Shimada, Kazuki +11 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Speech Corpus of Ainu Folklore and End-to-end Speech Recognition for Ainu Language
2020/02/16 by Kohei Matsuura, Matsuura, Kohei, Sei Ueno +7 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Time-domain Speech Enhancement Assisted by Multi-resolution Frequency Encoder and Decoder
2023/03/26 by Hao Shi, Masato Mimura, Shi, Hao +7 · 2 citations
Computer Science · Engineering · #Advanced Adaptive Filtering Techniques #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Multilingual End-to-End Speech Translation
2019/10/01 by Hirofumi Inaguma, Inaguma, Hirofumi, Kevin Duh +5 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
- Multilingual Turn-taking Prediction Using Voice Activity Projection
2024/03/11 by Koji Inoue, Bing’er Jiang, Inoue, Koji +7 · 3 citations
Computer Science · #Speech and dialogue systems
- Designing Precise and Robust Dialogue Response Evaluators
2020/04/10 by Tianyu Zhao, Zhao, Tianyu, Divesh Lala +3 · 1 citation
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Speech and dialogue systems #Topic Modeling
- Alignment Knowledge Distillation for Online Streaming Attention-based Speech Recognition
2021/02/28 by Hirofumi Inaguma, Tatsuya Kawahara, Inaguma, Hirofumi +1 · 1 citation
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
- Non-autoregressive End-to-end Speech Translation with Parallel Autoregressive Rescoring
2021/09/09 by Hirofumi Inaguma, Yosuke Higuchi, Inaguma, Hirofumi +7 · 1 citation
Computer Science · #Natural Language Processing Techniques #Topic Modeling #Speech Recognition and Synthesis
- Reasoning before Responding: Integrating Commonsense-based Causality Explanation for Empathetic Response Generation
2023/07/28 by Yahui Fu, Koji Inoue, Fu, Yahui +5 · 2 citations
Computer Science · Medicine · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Simulation-Based Education in Healthcare #Topic Modeling
- End-to-end Speech-to-Punctuated-Text Recognition
2022/07/07 by Jumon Nozaki, Nozaki, Jumon, Tatsuya Kawahara +5 · 1 citation
Computer Science · #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling
- Yeah, Un, Oh: Continuous and Real-time Backchannel Prediction with Fine-tuning of Voice Activity Projection
2024/10/21 by Koji Inoue, Inoue, Koji, Divesh Lala +5 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Human-Computer Interaction (cs.HC) #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- A Noise-Robust Turn-Taking System for Real-World Dialogue Robots: A Field Experiment
2025/03/08 by Koji Inoue, Yuki Okafuji, Inoue, Koji +9 · 1 citation
Computer Science · Engineering · Psychology · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Robotics (cs.RO) #Robotics and Automated Systems #Social Robot Interaction and HRI #Sound (cs.SD) #Speech and dialogue systems
- Enhancing Long-term RAG Chatbots with Psychological Models of Memory Importance and Forgetting
2024/09/19 by Ryuichi Sumida, Sumida, Ryuichi, Koji Inoue +3 · 1 citation
Computer Science · Medicine · #AI in Service Interactions #Artificial Intelligence in Healthcare and Education
- Human-Like Embodied AI Interviewer: Employing Android ERICA in Real\n International Conference
2024/12/13 by Zi Haur Pang, Yahui Fu, Pang, Zi Haur +9 · 1 citation
Engineering · #Robotics and Automated Systems
- Serialized Speech Information Guidance with Overlapped Encoding Separation for Multi-Speaker Automatic Speech Recognition
2024/09/01 by Hao Shi, Shi, Hao, Yuan Gao +5 · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Exploration of Adapter for Noise Robust Automatic Speech Recognition
2024/02/28 by Hao Shi, Shi, Hao, Tatsuya Kawahara +1 · 1 citation
Computer Science · #Speech Recognition and Synthesis
- Improving Zero-Shot Phonetic Classification through Language-Agnostic Articulatory Features
2026/07/26 by Ryo Magoshi, Jaeyoung Lee, Shinsuke Sakai +1
#cs.SD #eess.AS
- Memory-Driven Self-Disclosure and Relational Turning Points: A Longitudinal Multimodal Study of Human-AI Interaction
2026/07/16 by Ryuichi Sumida, Mao Saeki, Masaki Eguchi +4
Psychology · Computer Science · #Social Robot Interaction and HRI #Digital Mental Health Interventions #AI in Service Interactions
- On the Structure of Address in Multi-Party Dialogue: From Discrete Labels to Continuous Levels
2026/07/17 by Taiga Mori, Koji Inoue, Divesh Lala +1
#cs.CL #cs.AI #cs.HC