vix.ing · top · new · best · stats · spec

Ariya Rastrow

  1. Just ASK: Building an Architecture for Extensible Self-Service Spoken Language Understanding
    2017/11/01 by Anjishnu Kumar, Arpit Gupta, Kumar, Anjishnu +22 · 6 citations
    Computer Science · #68T50 #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Neural and Evolutionary Computing (cs.NE) #Software Engineering (cs.SE) #Speech and dialogue systems #Topic Modeling #cs.AI #cs.CL #cs.NE #cs.SE #msc:68T50
  2. Streaming Language Identification using Combination of Acoustic Representations and ASR Hypotheses
    2020/06/01 by Chander Chandak, Zeynab Raeesy, Chandak, Chander +13 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Authorship Attribution and Profiling #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  3. Towards ASR Robust Spoken Language Understanding Through In-Context Learning With Word Confusion Networks
    2024/01/05 by Kevin Everson, Everson, Kevin, Yile Gu +23 · 3 citations
    Computer Science · #Speech and dialogue systems #Topic Modeling #Natural Language Processing Techniques
  4. Personalized Predictive ASR for Latency Reduction in Voice Assistants
    2023/05/23 by Andreas Schwarz, Schwarz, Andreas, Di He +7 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
  5. Device-directed Utterance Detection
    2018/08/07 by Sri Harish Mallidi, Roland Maas, Mallidi, Sri Harish +9 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  6. Neural Machine Translation for Multilingual Grapheme-to-Phoneme\n Conversion
    2020/06/25 by Alex Sokolov, Sokolov, Alex, Tracy Rohlin +3 · 1 citation
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and dialogue systems #Topic Modeling
  7. Streaming End-to-End Bilingual ASR Systems with Joint Language Identification
    2020/07/08 by Surabhi Punjabi, Harish Arsikere, Punjabi, Surabhi +25 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
  8. Compute Cost Amortized Transformer for Streaming ASR
    2022/07/05 by Yi Xie, Jonathan J. Macoskey, Xie, Yi +13 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  9. Speech Recognition Rescoring with Large Speech-Text Foundation Models
    2024/09/25 by Prashanth Gurunath Shivakumar, Shivakumar, Prashanth Gurunath, Jari Kolehmainen +11 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  10. Using External Off-Policy Speech-To-Text Mappings in Contextual End-To-End Automated Speech Recognition
    2023/01/06 by David M. Chan, Chan, David M., Shalini Ghosh +5 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  11. Lookahead When It Matters: Adaptive Non-causal Transformers for Streaming Neural Transducers
    2023/05/07 by Grant P. Strimel, Yi Xie, Strimel, Grant P. +9 · 1 citation
    Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
  12. Federated Self-Learning with Weak Supervision for Speech Recognition
    2023/06/21 by Milind Rao, Rao, Milind, Gopinath Chennupati +11 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  13. Scaling Laws for Discriminative Speech Recognition Rescoring Models
    2023/06/27 by Yile Gu, Prashanth Gurunath Shivakumar, Gu, Yile +9 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  14. Discriminative Speech Recognition Rescoring with Pre-trained Language Models
    2023/10/10 by Prashanth Gurunath Shivakumar, Shivakumar, Prashanth Gurunath, Jari Kolehmainen +9 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  15. Multi-Modal Retrieval For Large Language Model Based Speech Recognition
    2024/06/13 by Jari Kolehmainen, Aditya Gourav, Kolehmainen, Jari +13 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Information Retrieval (cs.IR) #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
  16. CA-SSLR: Condition-Aware Self-Supervised Learning Representation for Generalized Speech Processing
    2024/12/05 by Yen-Ju Lu, Lu, Yen-Ju, Jing Liu +11 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering