Tucker, George
- Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
2025/07/07 by Gheorghe Comanici, Eric Bieber, Comanici, Gheorghe +6844 · 8 voices · 1374 citations
#cs.CL #cs.AI
- Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
2024/03/08 by Gemini Robotics Team, Petko Georgiev, Gemini Team +2277 · 4 voices · 559 citations
Computer Science · #Semantic Web and Ontologies
- Model-Based Reinforcement Learning for Atari
2019/03/01 by Lukasz Kaiser, Łukasz Kaiser, Mohammad Babaeizadeh +29 · 3 voices · 59 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence in Games #Reinforcement Learning in Robotics #cs.LG #stat.ML
- Training Language Models to Self-Correct via Reinforcement Learning
2024/09/19 by Aviral Kumar, Kumar, Aviral, Vincent Zhuang +36 · 2 voices · 76 citations
Computer Science · #Speech and dialogue systems #Topic Modeling #Natural Language Processing Techniques
- Conservative Q-Learning for Offline Reinforcement Learning
2020/06/08 by Kumar, Aviral, Zhou, Aurick, Tucker, George +1 · 172 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
2020/05/04 by Sergey Levine, Aviral Kumar, Levine, Sergey +5 · 203 citations
Computer Science · Engineering · #Reinforcement Learning in Robotics #Evolutionary Algorithms and Applications #Scheduling and Optimization Algorithms
- D4RL: Datasets for Deep Data-Driven Reinforcement Learning
2020/04/15 by Justin Fu, Fu, Justin, Aviral Kumar +7 · 156 citations
Computer Science · Engineering · #Reinforcement Learning in Robotics #Data Stream Mining Techniques #Autonomous Vehicle Technology and Safety
- Soft Actor-Critic Algorithms and Applications
2018/12/13 by Tuomas Haarnoja, Aurick Zhou, Haarnoja, Tuomas +19 · 109 citations
Computer Science · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Games #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Robotics (cs.RO)
- Gemma: Open Models Based on Gemini Research and Technology
2024/03/13 by Gemma Team, Thomas Mésnard, Cassidy Hardin +203 · 148 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Multi-Agent Systems and Negotiation
- Stabilizing Off-Policy Q-Learning via Bootstrapping Error Reduction
2019/06/03 by Aviral Kumar, Justin Fu, Kumar, Aviral +5 · 61 citations
Computer Science · Engineering · #Reinforcement Learning in Robotics #Adaptive Dynamic Programming Control #Smart Grid Energy Management
- On Variational Bounds of Mutual Information
2019/05/16 by Ben Poole, Poole, Ben, Sherjil Ozair +7 · 56 citations
Computer Science · #Advanced Neural Network Applications #FOS: Computer and information sciences #Face and Expression Recognition #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and ELM
- Behavior Regularized Offline Reinforcement Learning
2019/11/26 by Yifan Wu, Wu, Yifan, George Tucker +3 · 57 citations
Computer Science · Decision Sciences · Engineering · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Smart Grid Energy Management
- Regularizing Neural Networks by Penalizing Confident Output Distributions
2017/01/23 by Gabriel Pereyra, Pereyra, Gabriel, George Tucker +7 · 30 citations
Computer Science · #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Data Classification #Neural Networks and Applications #Neural and Evolutionary Computing (cs.NE)
- Waymax: An Accelerated, Data-Driven Simulator for Large-Scale Autonomous Driving Research
2023/10/12 by Cole Gulino, Justin Fu, Gulino, Cole +41 · 24 citations
Engineering · Psychology · #Autonomous Vehicle Technology and Safety #Traffic control and management #Human-Automation Interaction and Safety
- Imitation Is Not Enough: Robustifying Imitation with Reinforcement Learning for Challenging Driving Scenarios
2022/12/21 by Yiren Lu, Justin Fu, Lu, Yiren +21 · 19 citations
Energy · Engineering · #Artificial Intelligence (cs.AI) #Autonomous Vehicle Technology and Safety #Energy, Environment, and Transportation Policies #FOS: Computer and information sciences #I.2.6 #I.2.9 #Robotics (cs.RO) #Traffic and Road Safety
- The Laplacian in RL: Learning Representations with Efficient Approximations
2018/10/10 by Wu, Yifan, Tucker, George, Nachum, Ofir · 12 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Learning to Walk via Deep Reinforcement Learning
2018/12/26 by Tuomas Haarnoja, Sehoon Ha, Haarnoja, Tuomas +9 · 13 citations
Engineering · Computer Science · #Robotic Locomotion and Control #Prosthetics and Rehabilitation Robotics #Reinforcement Learning in Robotics
- Don't Blame the ELBO! A Linear VAE Perspective on Posterior Collapse
2019/11/06 by James Lucas, George Tucker, Lucas, James +5 · 11 citations
Computer Science · Physics and Astronomy · #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Image and Signal Denoising Methods #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Model Reduction and Neural Networks
- Offline Q-Learning on Diverse Multi-Task Data Both Scales And Generalizes
2022/11/28 by Kumar, Aviral, Agarwal, Rishabh, Geng, Xinyang +2 · 13 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG)
- Gemini: A Family of Highly Capable Multimodal Models
2023/12/19 by Gemini Robotics Team, Rohan Anil, Gemini Team +2692 · 9 voices
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Natural Language Processing Techniques #cs.AI #cs.CL #cs.CV
- Meta-Learning without Memorization
2019/12/09 by Mingzhang Yin, George Tucker, Yin, Mingzhang +7 · 7 citations
Computer Science · #Domain Adaptation and Few-Shot Learning #Multimodal Machine Learning Applications #Advanced Neural Network Applications
- REBAR: Low-variance, unbiased gradient estimates for discrete latent\n variable models
2017/03/21 by George Tucker, Andriy Mnih, Tucker, George +7 · 4 citations
Computer Science · #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning in Healthcare #Topic Modeling
- Filtering Variational Objectives
2017/05/25 by Maddison, Chris J., Lawson, Dieterich, Tucker, George +5 · 3 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Neural and Evolutionary Computing (cs.NE)
- DR3: Value-Based Deep Reinforcement Learning Requires Explicit Regularization
2021/12/09 by Aviral Kumar, Rishabh Agarwal, Kumar, Aviral +9 · 5 citations
Computer Science · #Reinforcement Learning in Robotics #Domain Adaptation and Few-Shot Learning #Machine Learning and ELM
- Deep Bayesian Bandits Showdown: An Empirical Comparison of Bayesian Deep Networks for Thompson Sampling
2018/02/25 by Carlos Riquelme, George Tucker, Riquelme, Carlos +3 · 3 citations
Decision Sciences · Computer Science · #Advanced Bandit Algorithms Research #Data Stream Mining Techniques #Gaussian Processes and Bayesian Inference
- The Mirage of Action-Dependent Baselines in Reinforcement Learning
2018/02/27 by Tucker, George, Bhupatiraju, Surya, Gu, Shixiang +3 · 1 citation
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Smoothed Action Value Functions for Learning Gaussian Policies
2018/03/06 by Nachum, Ofir, Norouzi, Mohammad, Tucker, George +1 · 1 citation
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Doubly Reparameterized Gradient Estimators for Monte Carlo Objectives
2018/10/09 by Tucker, George, Lawson, Dieterich, Gu, Shixiang +1 · 1 citation
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Energy-Inspired Models: Learning with Sampler-Induced Distributions
2019/10/31 by Lawson, Dieterich, Tucker, George, Dai, Bo +1 · 1 citation
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Autoregressive Dynamics Models for Offline Policy Evaluation and Optimization
2021/04/28 by Zhang, Michael R., Paine, Tom Le, Nachum, Ofir +4 · 1 citation
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Oracle Inequalities for Model Selection in Offline Reinforcement Learning
2022/11/03 by Jonathan N. Lee, Lee, Jonathan N., George Tucker +7 · 1 citation
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Data Classification #Reinforcement Learning in Robotics
- Particle Value Functions
2017/03/16 by Chris J. Maddison, Dieterich Lawson, Maddison, Chris J. +11 · 1 citation
Economics, Econometrics and Finance · Decision Sciences · #Complex Systems and Time Series Analysis #Economic theories and models #Game Theory and Applications