Ariya Rastrow
- Just ASK: Building an Architecture for Extensible Self-Service Spoken Language Understanding
2017/11/01 by Anjishnu Kumar, Arpit Gupta, Kumar, Anjishnu +22 · 6 citations
Computer Science · #68T50 #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Neural and Evolutionary Computing (cs.NE) #Software Engineering (cs.SE) #Speech and dialogue systems #Topic Modeling #cs.AI #cs.CL #cs.NE #cs.SE #msc:68T50
- Streaming Language Identification using Combination of Acoustic Representations and ASR Hypotheses
2020/06/01 by Chander Chandak, Zeynab Raeesy, Chandak, Chander +13 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Authorship Attribution and Profiling #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Towards ASR Robust Spoken Language Understanding Through In-Context Learning With Word Confusion Networks
2024/01/05 by Kevin Everson, Everson, Kevin, Yile Gu +23 · 3 citations
Computer Science · #Speech and dialogue systems #Topic Modeling #Natural Language Processing Techniques
- Personalized Predictive ASR for Latency Reduction in Voice Assistants
2023/05/23 by Andreas Schwarz, Schwarz, Andreas, Di He +7 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
- Device-directed Utterance Detection
2018/08/07 by Sri Harish Mallidi, Roland Maas, Mallidi, Sri Harish +9 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Neural Machine Translation for Multilingual Grapheme-to-Phoneme\n Conversion
2020/06/25 by Alex Sokolov, Sokolov, Alex, Tracy Rohlin +3 · 1 citation
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and dialogue systems #Topic Modeling
- Streaming End-to-End Bilingual ASR Systems with Joint Language Identification
2020/07/08 by Surabhi Punjabi, Harish Arsikere, Punjabi, Surabhi +25 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
- Compute Cost Amortized Transformer for Streaming ASR
2022/07/05 by Yi Xie, Jonathan J. Macoskey, Xie, Yi +13 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Speech Recognition Rescoring with Large Speech-Text Foundation Models
2024/09/25 by Prashanth Gurunath Shivakumar, Shivakumar, Prashanth Gurunath, Jari Kolehmainen +11 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Using External Off-Policy Speech-To-Text Mappings in Contextual End-To-End Automated Speech Recognition
2023/01/06 by David M. Chan, Chan, David M., Shalini Ghosh +5 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Lookahead When It Matters: Adaptive Non-causal Transformers for Streaming Neural Transducers
2023/05/07 by Grant P. Strimel, Yi Xie, Strimel, Grant P. +9 · 1 citation
Computer Science · #Speech Recognition and Synthesis #Speech and Audio Processing #Music and Audio Processing
- Federated Self-Learning with Weak Supervision for Speech Recognition
2023/06/21 by Milind Rao, Rao, Milind, Gopinath Chennupati +11 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Scaling Laws for Discriminative Speech Recognition Rescoring Models
2023/06/27 by Yile Gu, Prashanth Gurunath Shivakumar, Gu, Yile +9 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Discriminative Speech Recognition Rescoring with Pre-trained Language Models
2023/10/10 by Prashanth Gurunath Shivakumar, Shivakumar, Prashanth Gurunath, Jari Kolehmainen +9 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Electrical engineering #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Multi-Modal Retrieval For Large Language Model Based Speech Recognition
2024/06/13 by Jari Kolehmainen, Aditya Gourav, Kolehmainen, Jari +13 · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Information Retrieval (cs.IR) #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- CA-SSLR: Condition-Aware Self-Supervised Learning Representation for Generalized Speech Processing
2024/12/05 by Yen-Ju Lu, Lu, Yen-Ju, Jing Liu +11 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering