vix.ing · top · new · best · stats · spec

Ivan Bulyko

  1. Paralinguistics-Enhanced Large Language Modeling of Spoken Dialogue
    2023/12/23 by Guan-Ting Lin, Prashanth Gurunath Shivakumar, Lin, Guan-Ting +15 · 17 citations
    Computer Science · #Topic Modeling #Sentiment Analysis and Opinion Mining #Natural Language Processing Techniques
  2. Align-SLM: Textless Spoken Language Models with Reinforcement Learning from AI Feedback
    2024/11/04 by Guan-Ting Lin, Lin, Guan-Ting, Prashanth Gurunath Shivakumar +11 · 13 citations
    Computer Science · #Topic Modeling #Natural Language Processing Techniques #Speech and dialogue systems
  3. Towards ASR Robust Spoken Language Understanding Through In-Context Learning With Word Confusion Networks
    2024/01/05 by Kevin Everson, Everson, Kevin, Yile Gu +23 · 3 citations
    Computer Science · #Speech and dialogue systems #Topic Modeling #Natural Language Processing Techniques
  4. Speech Recognition Rescoring with Large Speech-Text Foundation Models
    2024/09/25 by Prashanth Gurunath Shivakumar, Shivakumar, Prashanth Gurunath, Jari Kolehmainen +11 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  5. Scaling Laws for Discriminative Speech Recognition Rescoring Models
    2023/06/27 by Yile Gu, Prashanth Gurunath Shivakumar, Gu, Yile +9 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  6. Discriminative Speech Recognition Rescoring with Pre-trained Language Models
    2023/10/10 by Prashanth Gurunath Shivakumar, Jari Kolehmainen, Shivakumar, Prashanth Gurunath +9 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  7. Multi-Modal Retrieval For Large Language Model Based Speech Recognition
    2024/06/13 by Jari Kolehmainen, Kolehmainen, Jari, Aditya Gourav +13 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Information Retrieval (cs.IR) #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering