vix.ing · top · new · best · stats · spec

Karen Livescu

  1. Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
    2022/06/09 by Aarohi Srivastava, Abhinav Rastogi, Abhishek Rao +448 · 3 voices · 149 citations
    #cs.CL #cs.AI #cs.CY #cs.LG #stat.ML
  2. Comparative layer-wise analysis of self-supervised speech models
    2022/11/08 by Ankita Pasad, Pasad, Ankita, Bowen Shi +3 · 29 citations
    Computer Science · #Speech Recognition and Synthesis #Topic Modeling #Music and Audio Processing
  3. On Deep Multi-View Representation Learning: Objectives and Optimization
    2016/02/02 by Weiran Wang, Raman Arora, Wang, Weiran +5 · 13 citations
    Computer Science · #Advanced Image and Video Retrieval Techniques #FOS: Computer and information sciences #Image Retrieval and Classification Techniques #Machine Learning (cs.LG) #Video Analysis and Summarization
  4. Chess as a Testbed for Language Model State Tracking
    2021/02/26 by Shubham Toshniwal, Sam M. Wiseman, Toshniwal, Shubham +5 · 12 citations
    Computer Science · Economics, Econometrics and Finance · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Sports Analytics and Performance #Topic Modeling
  5. What Do Self-Supervised Speech Models Know About Words?
    2023/06/30 by Ankita Pasad, Chung-Ming Chien, Pasad, Ankita +5 · 15 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Speech Recognition and Synthesis #Speech and dialogue systems #Topic Modeling #electronic engineering #information engineering
  6. On The Landscape of Spoken Language Models: A Comprehensive Survey
    2025/04/11 by Siddhant Arora, Arora, Siddhant, Kai‐Wei Chang +17 · 40 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
  7. SLUE: New Benchmark Tasks for Spoken Language Understanding Evaluation\n on Natural Speech
    2021/11/19 by Suwon Shon, Ankita Pasad, Shon, Suwon +11 · 9 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Natural Language Processing Techniques #Sound (cs.SD) #Speech and dialogue systems #Topic Modeling #electronic engineering #information engineering
  8. A Comparison of Techniques for Language Model Integration in\n Encoder-Decoder Speech Recognition
    2018/07/27 by Shubham Toshniwal, Toshniwal, Shubham, Anjuli Kannan +9 · 6 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  9. Self-Supervised Speech Representations are More Phonetic than Semantic
    2024/06/12 by Kwanghee Choi, Choi, Kwanghee, Ankita Pasad +9 · 15 citations
    Computer Science · #Speech and dialogue systems #Speech Recognition and Synthesis #Natural Language Processing Techniques
  10. Towards Robust Speech Representation Learning for Thousands of Languages
    2024/06/30 by William Chen, Chen, William, Wangyou Zhang +17 · 16 citations
    Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques #Speech and dialogue systems
  11. Towards Universal Paraphrastic Sentence Embeddings
    2015/11/25 by John Wieting, Mohit Bansal, Wieting, John +5 · 5 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Topic Modeling
  12. Open-Domain Sign Language Translation Learned from Online Video
    2022/05/25 by Bowen Shi, Diane Brentari, Shi, Bowen +5 · 7 citations
    Computer Science · Psychology · #Hand Gesture Recognition Systems #Hearing Impairment and Communication #Human Pose and Action Recognition
  13. Learning to Ignore: Long Document Coreference with Bounded Memory Neural\n Networks
    2020/10/06 by Shubham Toshniwal, Toshniwal, Shubham, Sam M. Wiseman +7 · 5 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Semantic Web and Ontologies #Topic Modeling
  14. Approaching Deep Learning through the Spectral Dynamics of Weights
    2024/08/21 by David Yunis, Yunis, David, Kumar Kshitij Patel +14 · 2 voices · 7 citations
    Computer Science · #Advanced Neural Network Applications #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Stochastic Gradient Optimization Techniques #cs.AI #cs.LG
  15. Multi-view Recurrent Neural Acoustic Word Embeddings
    2016/11/14 by Wanjia He, He, Wanjia, Weiran Wang +3 · 4 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Music and Audio Processing #Natural Language Processing Techniques #Speech Recognition and Synthesis
  16. Discriminative Acoustic Word Embeddings: Recurrent Neural Network-Based Approaches
    2016/11/08 by Shane Settle, Settle, Shane, Karen Livescu +1 · 6 citations
    Computer Science · #Music and Audio Processing #Speech Recognition and Synthesis #Time Series Analysis and Forecasting
  17. Pre-training on high-resource speech recognition improves low-resource\n speech-to-text translation
    2018/09/05 by Sameer Bansal, Herman Kamper, Bansal, Sameer +7 · 3 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling
  18. An embedded segmental K-means model for unsupervised segmentation and\n clustering of speech
    2017/03/23 by Herman Kamper, Karen Livescu, Kamper, Herman +3 · 3 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and Audio Processing
  19. SignMusketeers: An Efficient Multi-Stream Approach for Sign Language Translation at Scale
    2024/06/11 by Shester Gueuwou, Gueuwou, Shester, Xiaodan Du +5 · 6 citations
    Computer Science · Psychology · Arts and Humanities · #Hand Gesture Recognition Systems #Hearing Impairment and Communication #Subtitles and Audiovisual Media
  20. SHuBERT: Self-Supervised Sign Language Representation Learning via Multi-Stream Cluster Prediction
    2024/11/25 by Shester Gueuwou, Xiaodan Du, Gueuwou, Shester +7 · 4 citations
    Computer Science · #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Hand Gesture Recognition Systems
  21. Deep convolutional acoustic word embeddings using word-pair side\n information
    2015/10/05 by Herman Kamper, Kamper, Herman, Weiran Wang +3 · 2 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing #Topic Modeling
  22. Query-by-Example Search with Discriminative Neural Acoustic Word\n Embeddings
    2017/06/12 by Shane Settle, Settle, Shane, Keith Levin +5 · 1 citation
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Music and Audio Processing #Speech Recognition and Synthesis #Time Series Analysis and Forecasting
  23. On the Effects of Heterogeneous Data Sources on Speech-to-Text Foundation Models
    2024/06/13 by Jinchuan Tian, Tian, Jinchuan, Yifan Peng +9 · 3 citations
    Computer Science · #Speech Recognition and Synthesis #Natural Language Processing Techniques
  24. On the Evaluation of Speech Foundation Models for Spoken Language Understanding
    2024/06/14 by Siddhant Arora, Ankita Pasad, Arora, Siddhant +21 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
  25. American Sign Language fingerspelling recognition in the wild
    2018/10/26 by Bowen Shi, Shi, Bowen, Aurora Martinez Del Rio +11 · 1 citation
    Computer Science · Psychology · #Hand Gesture Recognition Systems #Human Pose and Action Recognition #Hearing Impairment and Communication
  26. A Correspondence Variational Autoencoder for Unsupervised Acoustic Word Embeddings
    2020/12/03 by Puyuan Peng, Herman Kamper, Peng, Puyuan +3 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  27. Fingerspelling Detection in American Sign Language
    2021/04/03 by Bowen Shi, Shi, Bowen, Diane Brentari +5 · 1 citation
    Computer Science · Psychology · #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Hand Gesture Recognition Systems #Hearing Impairment and Communication #Human Pose and Action Recognition
  28. Hierarchical Multitask Learning for CTC-based Speech Recognition
    2018/07/17 by Kalpesh Krishna, Krishna, Kalpesh, Shubham Toshniwal +3 · 2 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Speech Recognition and Synthesis #Speech and Audio Processing
  29. DiscreteSLU: A Large Language Model with Self-Supervised Discrete Speech Units for Spoken Language Understanding
    2024/06/13 by Suwon Shon, Shon, Suwon, Kwangyoun Kim +9 · 2 citations
    Computer Science · #Speech and dialogue systems #Speech Recognition and Synthesis #Natural Language Processing Techniques
  30. Toward Joint Language Modeling for Speech Units and Text
    2023/10/12 by Ju-Chieh Chou, Chung-Ming Chien, Chou, Ju-Chieh +13 · 2 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #electronic engineering #information engineering
  31. Searching for fingerspelled content in American Sign Language
    2022/03/24 by Bowen Shi, Diane Brentari, Shi, Bowen +5 · 1 citation
    Computer Science · Psychology · #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Hand Gesture Recognition Systems #Hearing Impairment and Communication #Human Pose and Action Recognition
  32. Self-Supervised Video Transformers for Isolated Sign Language Recognition
    2023/09/02 by Marcelo Sandoval-Castañeda, Sandoval-Castaneda, Marcelo, Yanhong Li +7 · 1 citation
    Computer Science · Psychology · Engineering · #Hand Gesture Recognition Systems #Hearing Impairment and Communication #Gait Recognition and Analysis
  33. Chunk-Distilled Language Modeling
    2024/12/31 by Yanhong Li, Li, Yanhong, Karen Livescu +2 · 1 citation
    Computer Science · #Natural Language Processing Techniques #Topic Modeling
  34. Targeted Linguistic Analysis of Sign Language Models with Minimal Translation Pairs
    2026/04/29 by Serpil Karabüklü, Kanishka Misra, Shester Gueuwou +3 · 3 voices
    #cs.CL