vix.ing · top · new · best · stats · spec

Mnih, Volodymyr

  1. Playing Atari with Deep Reinforcement Learning
    2013/12/19 by Volodymyr Mnih, Mnih, Volodymyr, Koray Kavukcuoglu +11 · 5 voices · 322 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence in Games #Reinforcement Learning in Robotics #cs.LG
  2. Reinforcement Learning with Unsupervised Auxiliary Tasks
    2016/11/16 by Max Jaderberg, Jaderberg, Max, Volodymyr Mnih +12 · 3 voices · 63 citations
    Computer Science · Decision Sciences · #Reinforcement Learning in Robotics #Advanced Bandit Algorithms Research #Data Stream Mining Techniques
  3. Recurrent Models of Visual Attention
    2014/06/24 by Volodymyr Mnih, Nicolas Heess, Mnih, Volodymyr +5 · 1 voice · 12 citations
    Computer Science · Mathematics · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #cs.CV #cs.LG #stat.ML
  4. Policy Distillation
    2015/11/19 by Andrei A. Rusu, Rusu, Andrei A., Sergio Gómez Colmenarejo +15 · 38 citations
    Computer Science · Engineering · #Neural Networks and Reservoir Computing #CCD and CMOS Imaging Sensors #Visual Attention and Saliency Detection
  5. Using Fast Weights to Attend to the Recent Past
    2016/10/20 by Jimmy Ba, Ba, Jimmy, Geoffrey E. Hinton +7 · 17 citations
    Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Neural Networks and Applications #Neural and Evolutionary Computing (cs.NE)
  6. Sample Efficient Actor-Critic with Experience Replay
    2016/11/03 by Wang, Ziyu, Bapst, Victor, Heess, Nicolas +4 · 12 citations
    #FOS: Computer and information sciences #Machine Learning (cs.LG)
  7. In-context Reinforcement Learning with Algorithm Distillation
    2022/10/25 by Michael Laskin, Laskin, Michael, Luyu Wang +25 · 18 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and ELM #Reinforcement Learning in Robotics
  8. Fast Task Inference with Variational Intrinsic Successor Features
    2019/06/12 by Steven Hansen, Will Dabney, Hansen, Steven +9 · 12 citations
    Computer Science · #Reinforcement Learning in Robotics #Adversarial Robustness in Machine Learning #Domain Adaptation and Few-Shot Learning
  9. The Uncertainty Bellman Equation and Exploration
    2017/09/15 by Brendan O’Donoghue, O'Donoghue, Brendan, Ian Osband +5 · 10 citations
    Decision Sciences · Computer Science · Engineering · #Simulation Techniques and Applications #Reinforcement Learning in Robotics #Reservoir Engineering and Simulation Methods
  10. Learning values across many orders of magnitude
    2016/02/24 by van Hasselt, Hado, Guez, Arthur, Hessel, Matteo +2 · 6 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Neural and Evolutionary Computing (cs.NE)
  11. Learning by Playing - Solving Sparse Reward Tasks from Scratch
    2018/02/28 by Riedmiller, Martin, Hafner, Roland, Lampe, Thomas +6 · 5 citations
    #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Robotics (cs.RO)
  12. Massively Parallel Methods for Deep Reinforcement Learning
    2015/07/15 by Nair, Arun, Srinivasan, Praveen, Blackwell, Sam +11 · 4 citations
    #Artificial Intelligence (cs.AI) #Distributed #FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural and Evolutionary Computing (cs.NE) #Parallel #and Cluster Computing (cs.DC)
  13. Strategic Attentive Writer for Learning Macro-Actions
    2016/06/15 by Alexander -, Alexander, Volodymyr Mnih +12 · 9 citations
    Computer Science · Psychology · #Artificial Intelligence in Games #Teaching and Learning Programming #Educational Games and Gamification
  14. Unsupervised Learning of Object Keypoints for Perception and Control
    2019/06/19 by Tejas Kulkarni, Ankush Gupta, Kulkarni, Tejas +11 · 5 citations
    Computer Science · #Advanced Neural Network Applications #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
  15. ElasticTok: Adaptive Tokenization for Image and Video
    2024/10/10 by Wilson Yan, Volodymyr Mnih, Yan, Wilson +9 · 1 voice · 7 citations
    Computer Science · #cs.LG
  16. Unsupervised Control Through Non-Parametric Discriminative Rewards
    2018/11/28 by Warde-Farley, David, Van de Wiele, Tom, Kulkarni, Tejas +3 · 3 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  17. Conditional Restricted Boltzmann Machines for Structured Output\n Prediction
    2012/02/14 by Volodymyr Mnih, Hugo Larochelle, Mnih, Volodymyr +3 · 2 citations
    Computer Science · Physics and Astronomy · #Generative Adversarial Networks and Image Synthesis #Music and Audio Processing #Model Reduction and Neural Networks
  18. Q-Learning in enormous action spaces via amortized approximate maximization
    2020/01/22 by Tom Van de Wiele, David Warde-Farley, Van de Wiele, Tom +5 · 1 voice · 1 citation
    Computer Science · Mathematics · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Domain Adaptation and Few-Shot Learning #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Human Pose and Action Recognition #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #cs.AI #cs.LG #stat.ML
  19. LMAct: A Benchmark for In-Context Imitation Learning with Long Multimodal Demonstrations
    2024/12/02 by Anian Ruoss, Fabio Pardo, Ruoss, Anian +9 · 1 voice · 4 citations
    Computer Science · #Domain Adaptation and Few-Shot Learning #Human Pose and Action Recognition #Multimodal Machine Learning Applications #cs.AI #cs.LG
  20. Vision-Language Models as a Source of Rewards
    2023/12/14 by Baumli, Kate, Baveja, Satinder, Behbahani, Feryal +24 · 3 citations
    #FOS: Computer and information sciences #Machine Learning (cs.LG)
  21. Combining policy gradient and Q-learning
    2016/11/05 by O'Donoghue, Brendan, Munos, Remi, Kavukcuoglu, Koray +1 · 1 citation
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Optimization and Control (math.OC)
  22. Relative Variational Intrinsic Control
    2020/12/14 by Kate Baumli, David Warde-Farley, Baumli, Kate +5 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
  23. Discovering Diverse Nearly Optimal Policies with Successor Features
    2021/06/01 by Zahavy, Tom, O'Donoghue, Brendan, Barreto, Andre +3 · 1 citation
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  24. SIMA 2: A Generalist Embodied Agent for Virtual Worlds
    2025/12/04 by Adrian Bolton, SIMA team, Bolton, Adrian +135 · 3 voices · 1 citation
    Computer Science · Psychology · #Multimodal Machine Learning Applications #Reinforcement Learning in Robotics #Social Robot Interaction and HRI #cs.AI #cs.RO