vix.ing · top · new · best · stats · spec

Park, Taejin

  1. TitaNet: Neural Model for speaker representation with 1D Depth-wise separable convolutions and global context
    2021/10/08 by Nithin Rao Koluguri, Koluguri, Nithin Rao, Taejin Park +3 · 22 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  2. Sortformer: Seamless Integration of Speaker Diarization and ASR by Bridging Timestamps and Tokens
    2024/09/10 by Taejin Park, Ivan Medennikov, Park, Taejin +15 · 16 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
  3. The CHiME-8 DASR Challenge for Generalizable and Array Agnostic Distant Automatic Speech Recognition and Diarization
    2024/07/23 by Cornell, Samuele, Park, Taejin, Huang, Steve +6 · 10 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  4. Enhancing Anomaly Detection in Financial Markets with an LLM-based Multi-Agent Framework
    2024/03/28 by Taejin Park, Park, Taejin · 8 citations
    Decision Sciences · Economics, Econometrics and Finance · #Complex Systems and Time Series Analysis #FOS: Economics and business #Financial Markets and Investment Strategies #Risk Management (q-fin.RM) #Stock Market Forecasting Methods
  5. Large Language Model Based Generative Error Correction: A Challenge and Baselines for Speech Recognition, Speaker Tagging, and Emotion Recognition
    2024/09/15 by Chao-Han Huck Yang, Yang, Chao-Han Huck, Taejin Park +39 · 4 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  6. META-CAT: Speaker-Informed Speech Embeddings via Meta Information Concatenation for Multi-talker ASR
    2024/09/18 by Jinhan Wang, Weiqing Wang, Wang, Jinhan +17 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
  7. NEST: Self-supervised Fast Conformer as All-purpose Seasoning to Speech Processing Tasks
    2024/08/23 by Huang, He, Park, Taejin, Dhawan, Kunal +6 · 2 citations
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  8. Resource-Efficient Adaptation of Speech Foundation Models for Multi-Speaker ASR
    2024/09/02 by Wang, Weiqing, Dhawan, Kunal, Park, Taejin +6 · 1 citation
    #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  9. SPGISpeech 2.0: Transcribed multi-speaker financial audio for speaker-tagged transcription
    2025/08/07 by Grossman, Raymond, Park, Taejin, Dhawan, Kunal +6 · 1 citation
    #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  10. Streaming Sortformer: Speaker Cache-Based Online Speaker Diarization with Arrival-Time Ordering
    2025/07/24 by Ivan Medennikov, Tae‐Jin Park, Medennikov, Ivan +13 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering