Françoise Beaufays
- Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
2025/07/07 by Gheorghe Comanici, Comanici, Gheorghe, Eric Bieber +6844 · 8 voices · 1459 citations
Computer Science · #cs.CL #cs.AI
- Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages
2023/03/02 by Yu Zhang, Zhang, Yu, Wei Han +55 · 1 voice · 52 citations
Computer Science · Engineering · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Natural Language Processing Techniques #Sound (cs.SD) #Speech Recognition and Synthesis #Topic Modeling #cs.CL #cs.SD #eess.AS #electronic engineering #information engineering
- Long Short-Term Memory Based Recurrent Neural Network Architectures for Large Vocabulary Speech Recognition
2014/02/05 by Haşim Sak, Andrew Senior, Sak, Haşim +3 · 50 citations
Computer Science · Mathematics · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Handwritten Text Recognition Techniques #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Natural Language Processing Techniques #Neural and Evolutionary Computing (cs.NE) #Speech Recognition and Synthesis #cs.CL #cs.LG #cs.NE #stat.ML
- Applied Federated Learning: Improving Google Keyboard Query Suggestions
2018/12/07 by Timothy T. Yang, Yang, Timothy, Galen Andrew +13 · 20 citations
Computer Science · Social Sciences · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Mobile Crowdsensing and Crowdsourcing #Privacy, Security, and Data Protection #Privacy-Preserving Technologies in Data
- Federated Evaluation of On-device Personalization
2019/10/22 by Kangkang Wang, Wang, Kangkang, Rajiv Mathews +9 · 7 citations
Computer Science · Social Sciences · #FOS: Computer and information sciences #Human Mobility and Location-Based Analysis #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Privacy-Preserving Technologies in Data #Recommender Systems and Techniques
- Federated Learning for Emoji Prediction in a Mobile Keyboard
2019/06/11 by Rajiv Mathews, Ramaswamy, Swaroop, Mathews, Rajiv +4 · 5 citations
Computer Science · Social Sciences · #Child Development and Digital Technology #Computation and Language (cs.CL) #Digital Communication and Language #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Training Production Language Models without Memorizing User Data
2020/09/21 by Swaroop Ramaswamy, Ramaswamy, Swaroop, Om Thakkar +9 · 9 citations
Computer Science · #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Mobile Crowdsensing and Crowdsourcing #Privacy-Preserving Technologies in Data #Stochastic Gradient Optimization Techniques
- Large-scale ASR Domain Adaptation using Self- and Semi-supervised Learning
2021/10/01 by Dongseong Hwang, Hwang, Dongseong, Ananya Misra +17 · 5 citations
Biochemistry, Genetics and Molecular Biology · Computer Science · #Audio and Speech Processing (eess.AS) #Cancer-related molecular mechanisms research #Computation and Language (cs.CL) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Multimodal Machine Learning Applications #Sound (cs.SD) #electronic engineering #information engineering
- Low-rank Gradient Approximation For Memory-Efficient On-device Training\n of Deep Neural Network
2020/01/24 by Mary Gooneratne, Gooneratne, Mary, Khe Chai Sim +9 · 3 citations
Engineering · Physics and Astronomy · #Audio and Speech Processing (eess.AS) #Electromagnetic Scattering and Analysis #FOS: Computer and information sciences #FOS: Electrical engineering #Indoor and Outdoor Localization Technologies #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Sound (cs.SD) #Sparse and Compressive Sensing Techniques #electronic engineering #information engineering
- Writing Across the World's Languages: Deep Internationalization for\n Gboard, the Google Keyboard
2019/12/03 by Daan van Esch, van Esch, Daan, Elnaz Sarbar +17 · 2 voices
Computer Science · Social Sciences · #Mobile Agent-Based Network Management #Multimedia Communication and Technology #Speech and dialogue systems #cs.CL #cs.HC
- Fast Contextual Adaptation with Neural Associative Memory for On-Device Personalized Speech Recognition
2021/10/05 by Tsendsuren Munkhdalai, Munkhdalai, Tsendsuren, Khe Chai Sim +11 · 3 citations
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Neural and Evolutionary Computing (cs.NE) #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems #electronic engineering #information engineering
- Mixture-of-Expert Conformer for Streaming Multilingual ASR
2023/05/25 by Ke Hu, Bo Li, Hu, Ke +7 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- On-Device Personalization of Automatic Speech Recognition Models for Disordered Speech
2021/06/18 by Katrin Tomanek, Tomanek, Katrin, Françoise Beaufays +7 · 2 citations
Computer Science · Medicine · #Speech Recognition and Synthesis #Voice and Speech Disorders #Speech and Audio Processing
- Extracting Targeted Training Data from ASR Models, and How to Mitigate It
2022/04/18 by Ehsan Amid, Amid, Ehsan, Om Thakkar +7 · 2 citations
Computer Science · #Topic Modeling #Natural Language Processing Techniques #Speech Recognition and Synthesis
- Fast and Accurate Recurrent Neural Network Acoustic Models for Speech Recognition
2015/07/24 by Haşim Sak, Andrew Senior, Sak, Haşim +5 · 1 citation
Computer Science · Mathematics · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music and Audio Processing #Neural and Evolutionary Computing (cs.NE) #Speech Recognition and Synthesis #Speech and Audio Processing #cs.CL #cs.LG #cs.NE #stat.ML
- Personalized Speech recognition on mobile devices
2016/03/10 by Ian McGraw, Rohit Prabhavalkar, McGraw, Ian +22 · 1 citation
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #cs.CL #cs.LG #cs.SD
- Improving Speech Recognition for African American English With Audio Classification
2023/09/16 by Shefali Garg, Zhouyuan Huo, Garg, Shefali +25 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Mobile Keyboard Input Decoding with Finite-State Transducers
2017/04/13 by Tom Ouyang, Ouyang, Tom, David Rybach +5 · 1 citation
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Speech Recognition and Synthesis #Speech and Audio Processing #Speech and dialogue systems
- An Investigation Into On-device Personalization of End-to-end Automatic\n Speech Recognition Models
2019/09/14 by Khe Chai Sim, Sim, Khe Chai, Petr Zadrazil +3 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Federated Learning of N-gram Language Models
2019/10/08 by Mingqing Chen, Chen, Mingqing, Ananda Theertha Suresh +11 · 1 citation
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Internet Traffic Analysis and Secure E-voting #Machine Learning (cs.LG) #Privacy-Preserving Technologies in Data #Topic Modeling
- Extending Multilingual Speech Synthesis to 100+ Languages without Transcribed Data
2024/02/29 by Takaaki Saeki, Saeki, Takaaki, Gary Wang +19 · 2 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and dialogue systems #electronic engineering #information engineering
- Enabling On-Device Training of Speech Recognition Models with Federated Dropout
2021/10/07 by Dhruv Guliani, Guliani, Dhruv, Lillian Zhou +13 · 1 citation
Computer Science · Engineering · #68T10 #Distributed #FOS: Computer and information sciences #I.2.7 #Internet Traffic Analysis and Secure E-voting #Machine Learning (cs.LG) #Parallel #Privacy-Preserving Technologies in Data #Traffic Prediction and Management Techniques #and Cluster Computing (cs.DC)
- Revealing and Protecting Labels in Distributed Training
2021/10/31 by Trung Dang, Om Thakkar, Dang, Trung +9 · 1 citation
Computer Science · Engineering · #Adversarial Robustness in Machine Learning #Privacy-Preserving Technologies in Data #Geophysical Methods and Applications
- Online Model Compression for Federated Learning with Large Models
2022/05/06 by Tien-Ju Yang, Yang, Tien-Ju, Yonghui Xiao +9 · 1 citation
Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural Networks and Applications #Speech Recognition and Synthesis #Speech and Audio Processing
- Federated Pruning: Improving Neural Network Efficiency with Federated Learning
2022/09/14 by Rongmei Lin, Yonghui Xiao, Lin, Rongmei +11 · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Music and Audio Processing #Privacy-Preserving Technologies in Data #Speech Recognition and Synthesis
- Efficient Domain Adaptation for Speech Foundation Models
2023/02/03 by Bo Li, Dongseong Hwang, Li, Bo +19 · 1 citation
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering