vix.ing · top · new · best · stats · spec

David Silver

  1. Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
    2017/12/05 by David Silver, Thomas Hubert, Silver, David +23 · 5 voices · 200 citations
    Computer Science · #Artificial Intelligence in Games #Reinforcement Learning in Robotics #Video Analysis and Summarization
  2. Playing Atari with Deep Reinforcement Learning
    2013/12/19 by Volodymyr Mnih, Koray Kavukcuoglu, Mnih, Volodymyr +11 · 5 voices · 326 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence in Games #Reinforcement Learning in Robotics #cs.LG
  3. Mastering Atari, Go, chess and shogi by planning with a learned model
    2019/11/19 by Julian Schrittwieser, Ioannis Antonoglou, Thomas Hubert +9 · 4 voices · 150 citations
    Computer Science · #Artificial Intelligence in Games #Reinforcement Learning in Robotics #AI-based Problem Solving and Planning
  4. Asynchronous Methods for Deep Reinforcement Learning
    2016/02/04 by Volodymyr Mnih, Adrià Puigdomènech Badia, Mehdi Mirza +5 · 4 voices · 158 citations
    #cs.LG
  5. Mastering the game of Stratego with model-free multiagent reinforcement learning
    2022/06/30 by Julien Perolat, Julien Pérolat, Bart De Vylder +38 · 6 voices · 21 citations
    Computer Science · Decision Sciences · #Reinforcement Learning in Robotics #Artificial Intelligence in Games #Advanced Bandit Algorithms Research
  6. Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
    2025/07/07 by Gheorghe Comanici, Eric Bieber, Comanici, Gheorghe +6844 · 8 voices · 1350 citations
    #cs.CL #cs.AI
  7. Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
    2024/03/08 by Gemini Robotics Team, Petko Georgiev, Gemini Team +2277 · 4 voices · 538 citations
    Computer Science · #Semantic Web and Ontologies
  8. Reinforcement Learning with Unsupervised Auxiliary Tasks
    2016/11/16 by Max Jaderberg, Volodymyr Mnih, Jaderberg, Max +12 · 3 voices · 64 citations
    Computer Science · Decision Sciences · #Reinforcement Learning in Robotics #Advanced Bandit Algorithms Research #Data Stream Mining Techniques
  9. Highly accurate protein structure prediction with AlphaFold
    2021/07/15 by John Jumper, Richard Evans, Alexander Pritzel +31 · 1143 citations
    Biochemistry, Genetics and Molecular Biology · Materials Science · #Enzyme Structure and Function #Machine Learning in Bioinformatics #Protein Structure and Dynamics
  10. Deep Reinforcement Learning with Double Q-learning
    2015/09/22 by Hado van Hasselt, Arthur Guez, van Hasselt, Hado +3 · 1 voice · 158 citations
    Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #cs.LG
  11. FeUdal Networks for Hierarchical Reinforcement Learning
    2017/03/03 by Alexander Sasha Vezhnevets, Simon Osindero, Vezhnevets, Alexander Sasha +11 · 2 voices · 48 citations
    Computer Science · #cs.AI
  12. Continuous control with deep reinforcement learning
    2015/09/09 by Timothy Lillicrap, Lillicrap, Timothy P., Jonathan J. Hunt +13 · 430 citations
    Computer Science · Engineering · #Adaptive Dynamic Programming Control #Advanced Control Systems Optimization #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
  13. Decoupled Neural Interfaces using Synthetic Gradients
    2016/08/18 by Max Jaderberg, Wojciech Marian Czarnecki, Jaderberg, Max +11 · 2 voices · 19 citations
    Computer Science · Engineering · #cs.LG
  14. Prioritized Experience Replay
    2015/11/18 by Tom Schaul, Schaul, Tom, John Quan +5 · 145 citations
    Neuroscience · Engineering · Computer Science · #Neural dynamics and brain function #Advanced Memory and Neural Computing #Reinforcement Learning in Robotics
  15. StarCraft II: A New Challenge for Reinforcement Learning
    2017/08/16 by Oriol Vinyals, Timo Ewalds, Vinyals, Oriol +48 · 1 voice · 47 citations
    Computer Science · Social Sciences · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Games #Digital Games and Media #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #cs.AI #cs.LG
  16. Improved protein structure prediction using potentials from deep learning
    2020/01/15 by Andrew W. Senior, Andrew Senior, Richard Evans +18 · 77 citations
    Biochemistry, Genetics and Molecular Biology · Materials Science · #Enzyme Structure and Function #Plant biochemistry and biosynthesis #Protein Structure and Dynamics
  17. Unsupervised Predictive Memory in a Goal-Directed Agent
    2018/03/28 by Greg Wayne, Wayne, Greg, Chia-Chun Hung +50 · 1 voice · 6 citations
    Computer Science · Neuroscience · #Explainable Artificial Intelligence (XAI) #Neural dynamics and brain function #Reinforcement Learning in Robotics #cs.LG #stat.ML
  18. Rainbow: Combining Improvements in Deep Reinforcement Learning
    2017/10/06 by Matteo Hessel, Joseph Modayil, Hessel, Matteo +17 · 88 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
  19. Deep Reinforcement Learning from Self-Play in Imperfect-Information Games
    2016/03/03 by Johannes Heinrich, Heinrich, Johannes, David Silver +1 · 1 voice · 12 citations
    #cs.LG #cs.AI #cs.GT
  20. Implicit Quantile Networks for Distributional Reinforcement Learning
    2018/06/14 by Will Dabney, Dabney, Will, Georg Ostrovski +5 · 33 citations
    Computer Science · #Adaptive Dynamic Programming Control #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Neural Networks and Applications #Reinforcement Learning in Robotics
  21. Successor Features for Transfer in Reinforcement Learning
    2016/06/16 by André Sales Barreto, Barreto, André, Will Dabney +11 · 30 citations
    Computer Science · #Reinforcement Learning in Robotics #Adaptive Dynamic Programming Control #Evolutionary Algorithms and Applications
  22. A Unified Game-Theoretic Approach to Multiagent Reinforcement Learning
    2017/11/02 by Marc Lanctot, Lanctot, Marc, Vinicius Zambaldi +16 · 1 voice · 20 citations
    Computer Science · #Artificial Intelligence in Games #Evolutionary Algorithms and Applications #Reinforcement Learning in Robotics #cs.AI #cs.GT #cs.LG #cs.MA
  23. Discovering Reinforcement Learning Algorithms
    2020/07/17 by Junhyuk Oh, Oh, Junhyuk, Matteo Hessel +11 · 3 voices · 2 citations
    #cs.LG #cs.AI
  24. Доочистка биологически очищенных фенольных сточных вод методом коагуляции с использованием FeCl3·6H2O
    2009/01/01 by John Jumper, Richard Evans, Alexander Pritzel +35 · 13 citations
    Environmental Science · #Environmental and Industrial Safety
  25. Learning Continuous Control Policies by Stochastic Value Gradients
    2015/10/30 by Nicolas Heess, Heess, Nicolas, Greg Wayne +9 · 16 citations
    Computer Science · Engineering · #Advanced Control Systems Optimization #FOS: Computer and information sciences #Fault Detection and Control Systems #Machine Learning (cs.LG) #Neural and Evolutionary Computing (cs.NE) #Reinforcement Learning in Robotics
  26. Bayesian Optimization in AlphaGo
    2018/12/17 by Yutian Chen, Aja Huang, Chen, Yutian +11 · 1 voice · 3 citations
    Computer Science · Mathematics · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #cs.AI #cs.LG #stat.ML
  27. Imagination-Augmented Agents for Deep Reinforcement Learning
    2017/07/19 by Théophane Weber, Weber, Théophane, Sébastien Racanière +26 · 12 citations
    Computer Science · #Reinforcement Learning in Robotics
  28. Online and Offline Reinforcement Learning by Planning with a Learned Model
    2021/04/13 by Julian Schrittwieser, Thomas Hubert, Schrittwieser, Julian +9 · 11 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
  29. Gemini: A Family of Highly Capable Multimodal Models
    2023/12/19 by Gemini Team, Rohan Anil, Sebastian Borgeaud +2682 · 9 voices
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #cs.AI #cs.CL #cs.CV
  30. Universal Successor Features Approximators
    2018/12/18 by Diana Borsa, André Barreto, Borsa, Diana +13 · 7 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Data Classification #Reinforcement Learning in Robotics
  31. Learning and Planning in Complex Action Spaces
    2021/04/13 by Thomas Hubert, Hubert, Thomas, Julian Schrittwieser +9 · 9 citations
    Computer Science · #Artificial Intelligence in Games #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
  32. Meta-Gradient Reinforcement Learning
    2018/05/24 by Zhongwen Xu, Xu, Zhongwen, Hado van Hasselt +3 · 6 citations
    Computer Science · #Machine Learning and Data Classification #Reinforcement Learning in Robotics #Data Stream Mining Techniques
  33. The Value Equivalence Principle for Model-Based Reinforcement Learning
    2020/11/06 by Christopher Grimm, Grimm, Christopher, André Barreto +5 · 4 citations
    Computer Science · #Reinforcement Learning in Robotics #Adversarial Robustness in Machine Learning #Explainable Artificial Intelligence (XAI)
  34. Discovery of Useful Questions as Auxiliary Tasks
    2019/09/10 by Vivek Veeriah, Matteo Hessel, Veeriah, Vivek +15 · 4 citations
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural Networks and Reservoir Computing #Reinforcement Learning in Robotics
  35. DataRater: Meta-Learned Dataset Curation
    2025/05/23 by Dan A. Calian, Calian, Dan A., Gregory Farquhar +21 · 3 voices · 4 citations
    #stat.ML #cs.AI #cs.LG
  36. Move Evaluation in Go Using Deep Convolutional Neural Networks
    2014/12/20 by Chris J. Maddison, Maddison, Chris J., Aja Huang +5 · 1 voice
    Computer Science · Economics, Econometrics and Finance · Psychology · #Artificial Intelligence in Games #Educational Games and Gamification #Sports Analytics and Performance #cs.LG #cs.NE
  37. Ureteral Carcinoma in Situ at Radical Cystectomy: Does the Margin Matter?
    1997/09/01 by David A. Silver, David Silver, Nicholas Stroumbakis +3 · 1 citation
    Medicine · #Bladder and Urothelial Cancer Treatments #Urinary and Genital Oncology Studies #Urological Disorders and Treatments