David Silver
- Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
2017/12/05 by David Silver, Thomas Hubert, Silver, David +23 · 5 voices · 200 citations
Computer Science · #Artificial Intelligence in Games #Reinforcement Learning in Robotics #Video Analysis and Summarization
- Playing Atari with Deep Reinforcement Learning
2013/12/19 by Volodymyr Mnih, Koray Kavukcuoglu, Mnih, Volodymyr +11 · 5 voices · 326 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence in Games #Reinforcement Learning in Robotics #cs.LG
- Mastering Atari, Go, chess and shogi by planning with a learned model
2019/11/19 by Julian Schrittwieser, Ioannis Antonoglou, Thomas Hubert +9 · 4 voices · 150 citations
Computer Science · #Artificial Intelligence in Games #Reinforcement Learning in Robotics #AI-based Problem Solving and Planning
- Asynchronous Methods for Deep Reinforcement Learning
2016/02/04 by Volodymyr Mnih, Adrià Puigdomènech Badia, Mehdi Mirza +5 · 4 voices · 158 citations
#cs.LG
- Mastering the game of Stratego with model-free multiagent reinforcement learning
2022/06/30 by Julien Perolat, Julien Pérolat, Bart De Vylder +38 · 6 voices · 21 citations
Computer Science · Decision Sciences · #Reinforcement Learning in Robotics #Artificial Intelligence in Games #Advanced Bandit Algorithms Research
- Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
2025/07/07 by Gheorghe Comanici, Eric Bieber, Comanici, Gheorghe +6844 · 8 voices · 1350 citations
#cs.CL #cs.AI
- Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
2024/03/08 by Gemini Robotics Team, Petko Georgiev, Gemini Team +2277 · 4 voices · 538 citations
Computer Science · #Semantic Web and Ontologies
- Reinforcement Learning with Unsupervised Auxiliary Tasks
2016/11/16 by Max Jaderberg, Volodymyr Mnih, Jaderberg, Max +12 · 3 voices · 64 citations
Computer Science · Decision Sciences · #Reinforcement Learning in Robotics #Advanced Bandit Algorithms Research #Data Stream Mining Techniques
- Highly accurate protein structure prediction with AlphaFold
2021/07/15 by John Jumper, Richard Evans, Alexander Pritzel +31 · 1143 citations
Biochemistry, Genetics and Molecular Biology · Materials Science · #Enzyme Structure and Function #Machine Learning in Bioinformatics #Protein Structure and Dynamics
- Deep Reinforcement Learning with Double Q-learning
2015/09/22 by Hado van Hasselt, Arthur Guez, van Hasselt, Hado +3 · 1 voice · 158 citations
Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #cs.LG
- FeUdal Networks for Hierarchical Reinforcement Learning
2017/03/03 by Alexander Sasha Vezhnevets, Simon Osindero, Vezhnevets, Alexander Sasha +11 · 2 voices · 48 citations
Computer Science · #cs.AI
- Continuous control with deep reinforcement learning
2015/09/09 by Timothy Lillicrap, Lillicrap, Timothy P., Jonathan J. Hunt +13 · 430 citations
Computer Science · Engineering · #Adaptive Dynamic Programming Control #Advanced Control Systems Optimization #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
- Decoupled Neural Interfaces using Synthetic Gradients
2016/08/18 by Max Jaderberg, Wojciech Marian Czarnecki, Jaderberg, Max +11 · 2 voices · 19 citations
Computer Science · Engineering · #cs.LG
- Prioritized Experience Replay
2015/11/18 by Tom Schaul, Schaul, Tom, John Quan +5 · 145 citations
Neuroscience · Engineering · Computer Science · #Neural dynamics and brain function #Advanced Memory and Neural Computing #Reinforcement Learning in Robotics
- StarCraft II: A New Challenge for Reinforcement Learning
2017/08/16 by Oriol Vinyals, Timo Ewalds, Vinyals, Oriol +48 · 1 voice · 47 citations
Computer Science · Social Sciences · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Games #Digital Games and Media #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #cs.AI #cs.LG
- Improved protein structure prediction using potentials from deep learning
2020/01/15 by Andrew W. Senior, Andrew Senior, Richard Evans +18 · 77 citations
Biochemistry, Genetics and Molecular Biology · Materials Science · #Enzyme Structure and Function #Plant biochemistry and biosynthesis #Protein Structure and Dynamics
- Unsupervised Predictive Memory in a Goal-Directed Agent
2018/03/28 by Greg Wayne, Wayne, Greg, Chia-Chun Hung +50 · 1 voice · 6 citations
Computer Science · Neuroscience · #Explainable Artificial Intelligence (XAI) #Neural dynamics and brain function #Reinforcement Learning in Robotics #cs.LG #stat.ML
- Rainbow: Combining Improvements in Deep Reinforcement Learning
2017/10/06 by Matteo Hessel, Joseph Modayil, Hessel, Matteo +17 · 88 citations
Computer Science · #Artificial Intelligence (cs.AI) #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
- Deep Reinforcement Learning from Self-Play in Imperfect-Information Games
2016/03/03 by Johannes Heinrich, Heinrich, Johannes, David Silver +1 · 1 voice · 12 citations
#cs.LG #cs.AI #cs.GT
- Implicit Quantile Networks for Distributional Reinforcement Learning
2018/06/14 by Will Dabney, Dabney, Will, Georg Ostrovski +5 · 33 citations
Computer Science · #Adaptive Dynamic Programming Control #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Neural Networks and Applications #Reinforcement Learning in Robotics
- Successor Features for Transfer in Reinforcement Learning
2016/06/16 by André Sales Barreto, Barreto, André, Will Dabney +11 · 30 citations
Computer Science · #Reinforcement Learning in Robotics #Adaptive Dynamic Programming Control #Evolutionary Algorithms and Applications
- A Unified Game-Theoretic Approach to Multiagent Reinforcement Learning
2017/11/02 by Marc Lanctot, Lanctot, Marc, Vinicius Zambaldi +16 · 1 voice · 20 citations
Computer Science · #Artificial Intelligence in Games #Evolutionary Algorithms and Applications #Reinforcement Learning in Robotics #cs.AI #cs.GT #cs.LG #cs.MA
- Discovering Reinforcement Learning Algorithms
2020/07/17 by Junhyuk Oh, Oh, Junhyuk, Matteo Hessel +11 · 3 voices · 2 citations
#cs.LG #cs.AI
- Доочистка биологически очищенных фенольных сточных вод методом коагуляции с использованием FeCl3·6H2O
2009/01/01 by John Jumper, Richard Evans, Alexander Pritzel +35 · 13 citations
Environmental Science · #Environmental and Industrial Safety
- Learning Continuous Control Policies by Stochastic Value Gradients
2015/10/30 by Nicolas Heess, Heess, Nicolas, Greg Wayne +9 · 16 citations
Computer Science · Engineering · #Advanced Control Systems Optimization #FOS: Computer and information sciences #Fault Detection and Control Systems #Machine Learning (cs.LG) #Neural and Evolutionary Computing (cs.NE) #Reinforcement Learning in Robotics
- Bayesian Optimization in AlphaGo
2018/12/17 by Yutian Chen, Aja Huang, Chen, Yutian +11 · 1 voice · 3 citations
Computer Science · Mathematics · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #cs.AI #cs.LG #stat.ML
- Imagination-Augmented Agents for Deep Reinforcement Learning
2017/07/19 by Théophane Weber, Weber, Théophane, Sébastien Racanière +26 · 12 citations
Computer Science · #Reinforcement Learning in Robotics
- Online and Offline Reinforcement Learning by Planning with a Learned Model
2021/04/13 by Julian Schrittwieser, Thomas Hubert, Schrittwieser, Julian +9 · 11 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
- Gemini: A Family of Highly Capable Multimodal Models
2023/12/19 by Gemini Team, Rohan Anil, Sebastian Borgeaud +2682 · 9 voices
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #cs.AI #cs.CL #cs.CV
- Universal Successor Features Approximators
2018/12/18 by Diana Borsa, André Barreto, Borsa, Diana +13 · 7 citations
Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Data Classification #Reinforcement Learning in Robotics
- Learning and Planning in Complex Action Spaces
2021/04/13 by Thomas Hubert, Hubert, Thomas, Julian Schrittwieser +9 · 9 citations
Computer Science · #Artificial Intelligence in Games #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
- Meta-Gradient Reinforcement Learning
2018/05/24 by Zhongwen Xu, Xu, Zhongwen, Hado van Hasselt +3 · 6 citations
Computer Science · #Machine Learning and Data Classification #Reinforcement Learning in Robotics #Data Stream Mining Techniques
- The Value Equivalence Principle for Model-Based Reinforcement Learning
2020/11/06 by Christopher Grimm, Grimm, Christopher, André Barreto +5 · 4 citations
Computer Science · #Reinforcement Learning in Robotics #Adversarial Robustness in Machine Learning #Explainable Artificial Intelligence (XAI)
- Discovery of Useful Questions as Auxiliary Tasks
2019/09/10 by Vivek Veeriah, Matteo Hessel, Veeriah, Vivek +15 · 4 citations
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural Networks and Reservoir Computing #Reinforcement Learning in Robotics
- DataRater: Meta-Learned Dataset Curation
2025/05/23 by Dan A. Calian, Calian, Dan A., Gregory Farquhar +21 · 3 voices · 4 citations
#stat.ML #cs.AI #cs.LG
- Move Evaluation in Go Using Deep Convolutional Neural Networks
2014/12/20 by Chris J. Maddison, Maddison, Chris J., Aja Huang +5 · 1 voice
Computer Science · Economics, Econometrics and Finance · Psychology · #Artificial Intelligence in Games #Educational Games and Gamification #Sports Analytics and Performance #cs.LG #cs.NE
- Ureteral Carcinoma in Situ at Radical Cystectomy: Does the Margin Matter?
1997/09/01 by David A. Silver, David Silver, Nicholas Stroumbakis +3 · 1 citation
Medicine · #Bladder and Urothelial Cancer Treatments #Urinary and Genital Oncology Studies #Urological Disorders and Treatments