vix.ing · top · new · best · stats · spec

Takahiro Shinozaki

  1. FlexMatch: Boosting Semi-Supervised Learning with Curriculum Pseudo Labeling
    2021/10/15 by Bowen Zhang, Yidong Wang, Zhang, Bowen +11 · 50 citations
    Computer Science · #Advanced Neural Network Applications #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Data Classification
  2. FreeMatch: Self-adaptive Thresholding for Semi-supervised Learning
    2022/05/15 by Yidong Wang, Wang, Yidong, Hao Chen +18 · 30 citations
    Computer Science · Medicine · #Domain Adaptation and Few-Shot Learning #Advanced Neural Network Applications #COVID-19 diagnosis using AI
  3. Deep Generic Representations for Domain-Generalized Anomalous Sound Detection
    2024/09/08 by Phurich Saengthong, Takahiro Shinozaki, Saengthong, Phurich +1 · 3 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  4. Exploiting Adapters for Cross-lingual Low-resource Speech Recognition
    2021/05/18 by Wenxin Hou, Zhu Han, Hou, Wenxin +11 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  5. Cross-domain Speech Recognition with Unsupervised Character-level Distribution Matching
    2021/04/15 by Wenxin Hou, Hou, Wenxin, Jindong Wang +7 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  6. Margin Calibration for Long-Tailed Visual Recognition
    2021/12/14 by Yidong Wang, Wang, Yidong, Bowen Zhang +9 · 1 citation
    Computer Science · #Advanced Neural Network Applications #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Data Classification
  7. Censer: Curriculum Semi-supervised Learning for Speech Recognition Based on Self-supervised Pre-training
    2022/06/16 by Bowen Zhang, Songjun Cao, Zhang, Bowen +9 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  8. Streaming Target-Speaker ASR with Neural Transducer
    2022/09/09 by Takafumi Moriya, Hiroshi Sato, Moriya, Takafumi +7 · 1 citation
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
  9. Self-Supervised Syllable Discovery Based on Speaker-Disentangled HuBERT
    2024/09/16 by Ryota Komatsu, Komatsu, Ryota, Takahiro Shinozaki +1 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
  10. MemNMF: Memory-Augmented NMF on LPC Spectra for Anomalous Sound Detection
    2026/07/24 by Phurich Saengthong, Takahiro Shinozaki
    #cs.SD #cs.LG