vix.ing · top · new · best · stats · spec

Tucker, George

  1. Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
    2025/07/07 by Gheorghe Comanici, Eric Bieber, Comanici, Gheorghe +6844 · 8 voices · 1374 citations
    #cs.CL #cs.AI
  2. Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
    2024/03/08 by Gemini Robotics Team, Petko Georgiev, Gemini Team +2277 · 4 voices · 559 citations
    Computer Science · #Semantic Web and Ontologies
  3. Model-Based Reinforcement Learning for Atari
    2019/03/01 by Lukasz Kaiser, Łukasz Kaiser, Mohammad Babaeizadeh +29 · 3 voices · 59 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence in Games #Reinforcement Learning in Robotics #cs.LG #stat.ML
  4. Training Language Models to Self-Correct via Reinforcement Learning
    2024/09/19 by Aviral Kumar, Kumar, Aviral, Vincent Zhuang +36 · 2 voices · 76 citations
    Computer Science · #Speech and dialogue systems #Topic Modeling #Natural Language Processing Techniques
  5. Conservative Q-Learning for Offline Reinforcement Learning
    2020/06/08 by Kumar, Aviral, Zhou, Aurick, Tucker, George +1 · 172 citations
    #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  6. Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
    2020/05/04 by Sergey Levine, Aviral Kumar, Levine, Sergey +5 · 203 citations
    Computer Science · Engineering · #Reinforcement Learning in Robotics #Evolutionary Algorithms and Applications #Scheduling and Optimization Algorithms
  7. D4RL: Datasets for Deep Data-Driven Reinforcement Learning
    2020/04/15 by Justin Fu, Fu, Justin, Aviral Kumar +7 · 156 citations
    Computer Science · Engineering · #Reinforcement Learning in Robotics #Data Stream Mining Techniques #Autonomous Vehicle Technology and Safety
  8. Soft Actor-Critic Algorithms and Applications
    2018/12/13 by Tuomas Haarnoja, Aurick Zhou, Haarnoja, Tuomas +19 · 109 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Games #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Robotics (cs.RO)
  9. Gemma: Open Models Based on Gemini Research and Technology
    2024/03/13 by Gemma Team, Thomas Mésnard, Cassidy Hardin +203 · 148 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Multi-Agent Systems and Negotiation
  10. Stabilizing Off-Policy Q-Learning via Bootstrapping Error Reduction
    2019/06/03 by Aviral Kumar, Justin Fu, Kumar, Aviral +5 · 61 citations
    Computer Science · Engineering · #Reinforcement Learning in Robotics #Adaptive Dynamic Programming Control #Smart Grid Energy Management
  11. On Variational Bounds of Mutual Information
    2019/05/16 by Ben Poole, Poole, Ben, Sherjil Ozair +7 · 56 citations
    Computer Science · #Advanced Neural Network Applications #FOS: Computer and information sciences #Face and Expression Recognition #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and ELM
  12. Behavior Regularized Offline Reinforcement Learning
    2019/11/26 by Yifan Wu, Wu, Yifan, George Tucker +3 · 57 citations
    Computer Science · Decision Sciences · Engineering · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Smart Grid Energy Management
  13. Regularizing Neural Networks by Penalizing Confident Output Distributions
    2017/01/23 by Gabriel Pereyra, Pereyra, Gabriel, George Tucker +7 · 30 citations
    Computer Science · #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Data Classification #Neural Networks and Applications #Neural and Evolutionary Computing (cs.NE)
  14. Waymax: An Accelerated, Data-Driven Simulator for Large-Scale Autonomous Driving Research
    2023/10/12 by Cole Gulino, Justin Fu, Gulino, Cole +41 · 24 citations
    Engineering · Psychology · #Autonomous Vehicle Technology and Safety #Traffic control and management #Human-Automation Interaction and Safety
  15. Imitation Is Not Enough: Robustifying Imitation with Reinforcement Learning for Challenging Driving Scenarios
    2022/12/21 by Yiren Lu, Justin Fu, Lu, Yiren +21 · 19 citations
    Energy · Engineering · #Artificial Intelligence (cs.AI) #Autonomous Vehicle Technology and Safety #Energy, Environment, and Transportation Policies #FOS: Computer and information sciences #I.2.6 #I.2.9 #Robotics (cs.RO) #Traffic and Road Safety
  16. The Laplacian in RL: Learning Representations with Efficient Approximations
    2018/10/10 by Wu, Yifan, Tucker, George, Nachum, Ofir · 12 citations
    #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  17. Learning to Walk via Deep Reinforcement Learning
    2018/12/26 by Tuomas Haarnoja, Sehoon Ha, Haarnoja, Tuomas +9 · 13 citations
    Engineering · Computer Science · #Robotic Locomotion and Control #Prosthetics and Rehabilitation Robotics #Reinforcement Learning in Robotics
  18. Don't Blame the ELBO! A Linear VAE Perspective on Posterior Collapse
    2019/11/06 by James Lucas, George Tucker, Lucas, James +5 · 11 citations
    Computer Science · Physics and Astronomy · #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Image and Signal Denoising Methods #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Model Reduction and Neural Networks
  19. Offline Q-Learning on Diverse Multi-Task Data Both Scales And Generalizes
    2022/11/28 by Kumar, Aviral, Agarwal, Rishabh, Geng, Xinyang +2 · 13 citations
    #FOS: Computer and information sciences #Machine Learning (cs.LG)
  20. Gemini: A Family of Highly Capable Multimodal Models
    2023/12/19 by Gemini Robotics Team, Rohan Anil, Gemini Team +2692 · 9 voices
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Natural Language Processing Techniques #cs.AI #cs.CL #cs.CV
  21. Meta-Learning without Memorization
    2019/12/09 by Mingzhang Yin, George Tucker, Yin, Mingzhang +7 · 7 citations
    Computer Science · #Domain Adaptation and Few-Shot Learning #Multimodal Machine Learning Applications #Advanced Neural Network Applications
  22. REBAR: Low-variance, unbiased gradient estimates for discrete latent\n variable models
    2017/03/21 by George Tucker, Andriy Mnih, Tucker, George +7 · 4 citations
    Computer Science · #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning in Healthcare #Topic Modeling
  23. Filtering Variational Objectives
    2017/05/25 by Maddison, Chris J., Lawson, Dieterich, Tucker, George +5 · 3 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Neural and Evolutionary Computing (cs.NE)
  24. DR3: Value-Based Deep Reinforcement Learning Requires Explicit Regularization
    2021/12/09 by Aviral Kumar, Rishabh Agarwal, Kumar, Aviral +9 · 5 citations
    Computer Science · #Reinforcement Learning in Robotics #Domain Adaptation and Few-Shot Learning #Machine Learning and ELM
  25. Deep Bayesian Bandits Showdown: An Empirical Comparison of Bayesian Deep Networks for Thompson Sampling
    2018/02/25 by Carlos Riquelme, George Tucker, Riquelme, Carlos +3 · 3 citations
    Decision Sciences · Computer Science · #Advanced Bandit Algorithms Research #Data Stream Mining Techniques #Gaussian Processes and Bayesian Inference
  26. The Mirage of Action-Dependent Baselines in Reinforcement Learning
    2018/02/27 by Tucker, George, Bhupatiraju, Surya, Gu, Shixiang +3 · 1 citation
    #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  27. Smoothed Action Value Functions for Learning Gaussian Policies
    2018/03/06 by Nachum, Ofir, Norouzi, Mohammad, Tucker, George +1 · 1 citation
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  28. Doubly Reparameterized Gradient Estimators for Monte Carlo Objectives
    2018/10/09 by Tucker, George, Lawson, Dieterich, Gu, Shixiang +1 · 1 citation
    #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  29. Energy-Inspired Models: Learning with Sampler-Induced Distributions
    2019/10/31 by Lawson, Dieterich, Tucker, George, Dai, Bo +1 · 1 citation
    #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  30. Autoregressive Dynamics Models for Offline Policy Evaluation and Optimization
    2021/04/28 by Zhang, Michael R., Paine, Tom Le, Nachum, Ofir +4 · 1 citation
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  31. Oracle Inequalities for Model Selection in Offline Reinforcement Learning
    2022/11/03 by Jonathan N. Lee, Lee, Jonathan N., George Tucker +7 · 1 citation
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Data Classification #Reinforcement Learning in Robotics
  32. Particle Value Functions
    2017/03/16 by Chris J. Maddison, Dieterich Lawson, Maddison, Chris J. +11 · 1 citation
    Economics, Econometrics and Finance · Decision Sciences · #Complex Systems and Time Series Analysis #Economic theories and models #Game Theory and Applications