Tomohiro Tanaka
- Deep versus Wide: An Analysis of Student Architectures for Task-Agnostic Knowledge Distillation of Self-Supervised Speech Models
2022/07/14 by Takanori Ashihara, Ashihara, Takanori, Takafumi Moriya +5 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- SpeechGLUE: How Well Can Self-Supervised Speech Models Capture Linguistic Knowledge?
2023/06/14 by Takanori Ashihara, Takafumi Moriya, Ashihara, Takanori +13 · 3 citations
Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Speech and dialogue systems
- Exploring Limits of Diffusion-Synthetic Training with Weakly Supervised Semantic Segmentation
2023/09/04 by Ryota Yoshihashi, Yoshihashi, Ryota, Yuya Otsuka +6 · 3 citations
Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Multimodal Machine Learning Applications
- Exploration of Language Dependency for Japanese Self-Supervised Speech Representation Models
2023/05/09 by Takanori Ashihara, Ashihara, Takanori, Takafumi Moriya +5 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Enrollment-less training for personalized voice activity detection
2021/06/23 by Naoki Makishima, Mana Ihori, Makishima, Naoki +9 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- End-to-End Joint Target and Non-Target Speakers ASR
2023/06/04 by Ryo Masumura, Masumura, Ryo, Naoki Makishima +27 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
- Mitochondrial dynamics in exercise physiology
2019/02/01 by Tomohiro Tanaka, Akiyuki Nishimura, Kazuhiro Nishiyama +4 · 1 citation
Biochemistry, Genetics and Molecular Biology · Medicine · #Mitochondrial Function and Pathology #Adipose Tissue and Metabolism #Autophagy in Disease and Therapy
- Improving Scheduled Sampling for Neural Transducer-based ASR
2023/05/25 by Takafumi Moriya, Takanori Ashihara, Moriya, Takafumi +9 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Transfer Learning from Pre-trained Language Models Improves End-to-End Speech Summarization
2023/06/07 by Kohei Matsuura, Takanori Ashihara, Matsuura, Kohei +11 · 1 citation
Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Topic Modeling