vix.ing · top · new · best · stats · spec

Tomohiro Tanaka

  1. Deep versus Wide: An Analysis of Student Architectures for Task-Agnostic Knowledge Distillation of Self-Supervised Speech Models
    2022/07/14 by Takanori Ashihara, Ashihara, Takanori, Takafumi Moriya +5 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
  2. SpeechGLUE: How Well Can Self-Supervised Speech Models Capture Linguistic Knowledge?
    2023/06/14 by Takanori Ashihara, Takafumi Moriya, Ashihara, Takanori +13 · 3 citations
    Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Speech and dialogue systems
  3. Exploring Limits of Diffusion-Synthetic Training with Weakly Supervised Semantic Segmentation
    2023/09/04 by Ryota Yoshihashi, Yoshihashi, Ryota, Yuya Otsuka +6 · 3 citations
    Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Multimodal Machine Learning Applications
  4. Exploration of Language Dependency for Japanese Self-Supervised Speech Representation Models
    2023/05/09 by Takanori Ashihara, Ashihara, Takanori, Takafumi Moriya +5 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  5. Enrollment-less training for personalized voice activity detection
    2021/06/23 by Naoki Makishima, Mana Ihori, Makishima, Naoki +9 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  6. End-to-End Joint Target and Non-Target Speakers ASR
    2023/06/04 by Ryo Masumura, Masumura, Ryo, Naoki Makishima +27 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
  7. Mitochondrial dynamics in exercise physiology
    2019/02/01 by Tomohiro Tanaka, Akiyuki Nishimura, Kazuhiro Nishiyama +4 · 1 citation
    Biochemistry, Genetics and Molecular Biology · Medicine · #Mitochondrial Function and Pathology #Adipose Tissue and Metabolism #Autophagy in Disease and Therapy
  8. Improving Scheduled Sampling for Neural Transducer-based ASR
    2023/05/25 by Takafumi Moriya, Takanori Ashihara, Moriya, Takafumi +9 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  9. Transfer Learning from Pre-trained Language Models Improves End-to-End Speech Summarization
    2023/06/07 by Kohei Matsuura, Takanori Ashihara, Matsuura, Kohei +11 · 1 citation
    Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Topic Modeling