vix.ing · top · new · best · stats · spec

Hoirin Kim

  1. FitHuBERT: Going Thinner and Deeper for Knowledge Distillation of Speech Self-Supervised Learning
    2022/07/01 by Yeonghyeon Lee, Kangwook Jang, Lee, Yeonghyeon +7 · 9 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
  2. Meta-Learning for Short Utterance Speaker Recognition with Imbalance Length Pairs
    2020/04/06 by Seong Min Kye, Kye, Seong Min, Youngmoon Jung +7 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  3. STaR: Distilling Speech Temporal Relation for Lightweight Speech Self-Supervised Learning Models
    2023/12/14 by Kangwook Jang, Jang, Kangwook, Sungnyun Kim +3 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  4. Dual Attention in Time and Frequency Domain for Voice Activity Detection
    2020/03/27 by Joohyung Lee, Lee, Joohyung, Youngmoon Jung +3 · 1 citation
    Computer Science · Engineering · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Indoor and Outdoor Localization Technologies #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  5. Multi-Task Network for Noise-Robust Keyword Spotting and Speaker Verification using CTC-based Soft VAD and Global Query Attention
    2020/05/08 by Myunghun Jung, Jung, Myunghun, Youngmoon Jung +5 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  6. Learning acoustic word embeddings with phonetically associated triplet network
    2018/11/07 by Hyungjun Lim, Lim, Hyungjun, Younggwan Kim +7 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Signal Processing (eess.SP) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  7. Asymmetric Proxy Loss for Multi-View Acoustic Word Embeddings
    2022/03/30 by Myunghun Jung, Jung, Myunghun, Hoirin Kim +1 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  8. Learning Video Temporal Dynamics with Cross-Modal Attention for Robust Audio-Visual Speech Recognition
    2024/07/04 by Sungnyun Kim, Kim, Sungnyun, Kangwook Jang +7 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Image and Signal Denoising Methods #Image and Video Processing (eess.IV) #Machine Learning (cs.LG) #Music and Audio Processing #Speech and Audio Processing #electronic engineering #information engineering
  9. AdaMS: Deep Metric Learning with Adaptive Margin and Adaptive Scale for Acoustic Word Discrimination
    2022/10/26 by Myunghun Jung, Jung, Myunghun, Hoirin Kim +1 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Information Retrieval (cs.IR) #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  10. Neural MOS Prediction for Synthesized Speech Using Multi-Task Learning With Spoofing Detection and Spoofing Type Classification
    2020/07/16 by Yeunju Choi, Choi, Yeunju, Youngmoon Jung +3 · 1 citation
    Computer Science · Medicine · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Voice and Speech Disorders #electronic engineering #information engineering