vix.ing · top · new · best · stats · spec

Françoise Beaufays

  1. Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
    2025/07/07 by Gheorghe Comanici, Comanici, Gheorghe, Eric Bieber +6844 · 8 voices · 1459 citations
    Computer Science · #cs.CL #cs.AI
  2. Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages
    2023/03/02 by Yu Zhang, Zhang, Yu, Wei Han +55 · 1 voice · 52 citations
    Computer Science · Engineering · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #cs.CL #cs.SD #eess.AS #electronic engineering #information engineering
  3. Long Short-Term Memory Based Recurrent Neural Network Architectures for Large Vocabulary Speech Recognition
    2014/02/05 by Haşim Sak, Andrew Senior, Sak, Haşim +3 · 50 citations
    Computer Science · Mathematics · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Handwritten Text Recognition Techniques #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Natural Language Processing Techniques #Neural and Evolutionary Computing (cs.NE) #Speech Recognition and Synthesis #cs.CL #cs.LG #cs.NE #stat.ML
  4. Applied Federated Learning: Improving Google Keyboard Query Suggestions
    2018/12/07 by Timothy T. Yang, Yang, Timothy, Galen Andrew +13 · 20 citations
    Computer Science · Social Sciences · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Mobile Crowdsensing and Crowdsourcing #Privacy, Security, and Data Protection #Privacy-Preserving Technologies in Data
  5. Federated Evaluation of On-device Personalization
    2019/10/22 by Kangkang Wang, Wang, Kangkang, Rajiv Mathews +9 · 7 citations
    Computer Science · Social Sciences · #FOS: Computer and information sciences #Human Mobility and Location-Based Analysis #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Privacy-Preserving Technologies in Data #Recommender Systems and Techniques
  6. Federated Learning for Emoji Prediction in a Mobile Keyboard
    2019/06/11 by Rajiv Mathews, Ramaswamy, Swaroop, Mathews, Rajiv +4 · 5 citations
    Computer Science · Social Sciences · #Child Development and Digital Technology #Computation and Language (cs.CL) #Digital Communication and Language #FOS: Computer and information sciences #Machine Learning (cs.LG)
  7. Training Production Language Models without Memorizing User Data
    2020/09/21 by Swaroop Ramaswamy, Ramaswamy, Swaroop, Om Thakkar +9 · 9 citations
    Computer Science · #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Mobile Crowdsensing and Crowdsourcing #Privacy-Preserving Technologies in Data #Stochastic Gradient Optimization Techniques
  8. Large-scale ASR Domain Adaptation using Self- and Semi-supervised Learning
    2021/10/01 by Dongseong Hwang, Hwang, Dongseong, Ananya Misra +17 · 5 citations
    Biochemistry, Genetics and Molecular Biology · Computer Science · #Audio and Speech Processing (eess.AS) #Cancer-related molecular mechanisms research #Computation and Language (cs.CL) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Multimodal Machine Learning Applications #Sound (cs.SD) #electronic engineering #information engineering
  9. Low-rank Gradient Approximation For Memory-Efficient On-device Training\n of Deep Neural Network
    2020/01/24 by Mary Gooneratne, Gooneratne, Mary, Khe Chai Sim +9 · 3 citations
    Engineering · Physics and Astronomy · #Audio and Speech Processing (eess.AS) #Electromagnetic Scattering and Analysis #FOS: Computer and information sciences #FOS: Electrical engineering #Indoor and Outdoor Localization Technologies #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Sound (cs.SD) #Sparse and Compressive Sensing Techniques #electronic engineering #information engineering
  10. Writing Across the World's Languages: Deep Internationalization for\n Gboard, the Google Keyboard
    2019/12/03 by Daan van Esch, van Esch, Daan, Elnaz Sarbar +17 · 2 voices
    Computer Science · Social Sciences · #Mobile Agent-Based Network Management #Multimedia Communication and Technology #Speech and dialogue systems #cs.CL #cs.HC
  11. Fast Contextual Adaptation with Neural Associative Memory for On-Device Personalized Speech Recognition
    2021/10/05 by Tsendsuren Munkhdalai, Munkhdalai, Tsendsuren, Khe Chai Sim +11 · 3 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Neural and Evolutionary Computing (cs.NE) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
  12. Mixture-of-Expert Conformer for Streaming Multilingual ASR
    2023/05/25 by Ke Hu, Bo Li, Hu, Ke +7 · 3 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  13. On-Device Personalization of Automatic Speech Recognition Models for Disordered Speech
    2021/06/18 by Katrin Tomanek, Tomanek, Katrin, Françoise Beaufays +7 · 2 citations
    Computer Science · Medicine · #Speech Recognition and Synthesis #Voice and Speech Disorders #Speech and Audio Processing
  14. Extracting Targeted Training Data from ASR Models, and How to Mitigate It
    2022/04/18 by Ehsan Amid, Amid, Ehsan, Om Thakkar +7 · 2 citations
    Computer Science · #Topic Modeling #Natural Language Processing Techniques #Speech Recognition and Synthesis
  15. Fast and Accurate Recurrent Neural Network Acoustic Models for Speech Recognition
    2015/07/24 by Haşim Sak, Andrew Senior, Sak, Haşim +5 · 1 citation
    Computer Science · Mathematics · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music and Audio Processing #Neural and Evolutionary Computing (cs.NE) #Speech Recognition and Synthesis #Speech and Audio Processing #cs.CL #cs.LG #cs.NE #stat.ML
  16. Personalized Speech recognition on mobile devices
    2016/03/10 by Ian McGraw, Rohit Prabhavalkar, McGraw, Ian +22 · 1 citation
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #cs.CL #cs.LG #cs.SD
  17. Improving Speech Recognition for African American English With Audio Classification
    2023/09/16 by Shefali Garg, Zhouyuan Huo, Garg, Shefali +25 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  18. Mobile Keyboard Input Decoding with Finite-State Transducers
    2017/04/13 by Tom Ouyang, Ouyang, Tom, David Rybach +5 · 1 citation
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems
  19. An Investigation Into On-device Personalization of End-to-end Automatic\n Speech Recognition Models
    2019/09/14 by Khe Chai Sim, Sim, Khe Chai, Petr Zadrazil +3 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
  20. Federated Learning of N-gram Language Models
    2019/10/08 by Mingqing Chen, Chen, Mingqing, Ananda Theertha Suresh +11 · 1 citation
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Internet Traffic Analysis and Secure E-voting #Machine Learning (cs.LG) #Privacy-Preserving Technologies in Data #Topic Modeling
  21. Extending Multilingual Speech Synthesis to 100+ Languages without Transcribed Data
    2024/02/29 by Takaaki Saeki, Saeki, Takaaki, Gary Wang +19 · 2 citations
    Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
  22. Enabling On-Device Training of Speech Recognition Models with Federated Dropout
    2021/10/07 by Dhruv Guliani, Guliani, Dhruv, Lillian Zhou +13 · 1 citation
    Computer Science · Engineering · #68T10 #Distributed #FOS: Computer and information sciences #I.2.7 #Internet Traffic Analysis and Secure E-voting #Machine Learning (cs.LG) #Parallel #Privacy-Preserving Technologies in Data #Traffic Prediction and Management Techniques #and Cluster Computing (cs.DC)
  23. Revealing and Protecting Labels in Distributed Training
    2021/10/31 by Trung Dang, Om Thakkar, Dang, Trung +9 · 1 citation
    Computer Science · Engineering · #Adversarial Robustness in Machine Learning #Privacy-Preserving Technologies in Data #Geophysical Methods and Applications
  24. Online Model Compression for Federated Learning with Large Models
    2022/05/06 by Tien-Ju Yang, Yang, Tien-Ju, Yonghui Xiao +9 · 1 citation
    Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural Networks and Applications #Speech Recognition and Synthesis #Speech and Audio Processing
  25. Federated Pruning: Improving Neural Network Efficiency with Federated Learning
    2022/09/14 by Rongmei Lin, Yonghui Xiao, Lin, Rongmei +11 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Music and Audio Processing #Privacy-Preserving Technologies in Data #Speech Recognition and Synthesis
  26. Efficient Domain Adaptation for Speech Foundation Models
    2023/02/03 by Bo Li, Dongseong Hwang, Li, Bo +19 · 1 citation
    Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering