vix.ing · top · new · best · stats · spec

Moriya, Takafumi

  1. SpeechGLUE: How Well Can Self-Supervised Speech Models Capture Linguistic Knowledge?
    2023/06/14 by Takanori Ashihara, Takafumi Moriya, Ashihara, Takanori +13 · 3 citations
    Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Speech and dialogue systems
  2. Deep versus Wide: An Analysis of Student Architectures for Task-Agnostic Knowledge Distillation of Self-Supervised Speech Models
    2022/07/14 by Ashihara, Takanori, Moriya, Takafumi, Matsuura, Kohei +1 · 2 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  3. SpeakerBeam-SS: Real-time Target Speaker Extraction with Lightweight Conv-TasNet and State Space Modeling
    2024/07/01 by Hiroshi Sato, Sato, Hiroshi, Takafumi Moriya +15 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Signal Processing (eess.SP) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  4. What Do Self-Supervised Speech and Speaker Models Learn? New Findings From a Cross Model Layer-Wise Analysis
    2024/01/31 by Takanori Ashihara, Ashihara, Takanori, Marc Delcroix +9 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  5. Cross-Modal Transformer-Based Neural Correction Models for Automatic Speech Recognition
    2021/07/04 by Tanaka, Tomohiro, Masumura, Ryo, Ihori, Mana +5 · 1 citation
    #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  6. End-to-End Joint Target and Non-Target Speakers ASR
    2023/06/04 by Ryo Masumura, Masumura, Ryo, Naoki Makishima +27 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
  7. Exploration of Language Dependency for Japanese Self-Supervised Speech Representation Models
    2023/05/09 by Ashihara, Takanori, Moriya, Takafumi, Matsuura, Kohei +1 · 1 citation
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  8. Improving Scheduled Sampling for Neural Transducer-based ASR
    2023/05/25 by Moriya, Takafumi, Ashihara, Takanori, Sato, Hiroshi +3 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #electronic engineering #information engineering
  9. Leveraging Large Text Corpora for End-to-End Speech Summarization
    2023/03/02 by Matsuura, Kohei, Ashihara, Takanori, Moriya, Takafumi +4 · 1 citation
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #electronic engineering #information engineering
  10. Sentence-wise Speech Summarization: Task, Datasets, and End-to-End Modeling with LM Knowledge Distillation
    2024/08/01 by Matsuura, Kohei, Ashihara, Takanori, Moriya, Takafumi +4 · 2 citations
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #electronic engineering #information engineering
  11. Transfer Learning from Pre-trained Language Models Improves End-to-End Speech Summarization
    2023/06/07 by Kohei Matsuura, Takanori Ashihara, Matsuura, Kohei +11 · 1 citation
    Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Topic Modeling
  12. Iterative Shallow Fusion of Backward Language Model for End-to-End Speech Recognition
    2023/10/17 by Atsunori Ogawa, Takafumi Moriya, Ogawa, Atsunori +7 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  13. Streaming Target-Speaker ASR with Neural Transducer
    2022/09/09 by Takafumi Moriya, Moriya, Takafumi, Hiroshi Sato +7 · 1 citation
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
  14. Alignment-Free Training for Transducer-based Multi-Talker ASR
    2024/09/30 by Takafumi Moriya, Shota Horiguchi, Moriya, Takafumi +13 · 2 citations
    Engineering · Computer Science · #Ultrasonics and Acoustic Wave Propagation #Fault Detection and Control Systems #Speech Recognition and Synthesis
  15. Factor-Conditioned Speaking-Style Captioning
    2024/06/27 by Ando, Atsushi, Moriya, Takafumi, Horiguchi, Shota +1 · 1 citation
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  16. Applying LLMs for Rescoring N-best ASR Hypotheses of Casual Conversations: Effects of Domain Adaptation and Context Carry-over
    2024/06/27 by Ogawa, Atsunori, Kamo, Naoyuki, Matsuura, Kohei +5 · 1 citation
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #electronic engineering #information engineering
  17. NTT Multi-Speaker ASR System for the DASR Task of CHiME-8 Challenge
    2024/09/09 by Naoyuki Kamo, Naohiro Tawara, Kamo, Naoyuki +33 · 2 citations
    Computer Science · Engineering · #Advanced Data Processing Techniques #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Robotics and Automated Systems #Speech Recognition and Synthesis #electronic engineering #information engineering