Anmol Gulati
- Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
2025/07/07 by Gheorghe Comanici, Eric Bieber, Comanici, Gheorghe +6844 · 8 voices · 1378 citations
#cs.CL #cs.AI
- Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
2024/03/08 by Gemini Robotics Team, Gemini Team, Petko Georgiev +2277 · 4 voices · 564 citations
Computer Science · #Semantic Web and Ontologies
- Conformer: Convolution-augmented Transformer for Speech Recognition
2020/05/16 by Anmol Gulati, Gulati, Anmol, James Qin +19 · 213 citations
Computer Science · #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- Conformer: Convolution-augmented Transformer for Speech Recognition
2020/10/25 by Anmol Gulati, James Qin, Chung‐Cheng Chiu +8 · 152 citations
Computer Science · #Music and Audio Processing #Speech Recognition and Synthesis #Speech and Audio Processing
- Is Grep All You Need? How Agent Harnesses Reshape Agentic Search
2026/05/14 by Sahil Sen, Akhil Kasturi, Elias Lumer +2 · 15 voices · 2 citations
#cs.CL
- Gemini: A Family of Highly Capable Multimodal Models
2023/12/19 by Gemini Robotics Team, Rohan Anil, Gemini Team +2692 · 9 voices · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Natural Language Processing Techniques #cs.AI #cs.CL #cs.CV
- ContextNet: Improving Convolutional Neural Networks for Automatic Speech Recognition with Global Context
2020/05/07 by Wei Han, Zhengdong Zhang, Han, Wei +15 · 3 citations
Computer Science · #Audio and Speech Processing (eess.AS) #Computation and Language (cs.CL) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #Speech and Audio Processing #electronic engineering #information engineering
- ScaleMCP: Dynamic and Auto-Synchronizing Model Context Protocol Tools for LLM Agents
2025/05/09 by Elias Lumer, Lumer, Elias, Anmol Gulati +7 · 14 citations
Computer Science · Social Sciences · #Artificial Intelligence in Law #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Topic Modeling