2019/10/09 by Louis Kirsch, Sjoerd van Steenkiste, Kirsch, Louis +3 · 19 citations
Computer Science · Mathematics · #Adaptive Dynamic Programming Control #Artificial Intelligence (cs.AI) #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #I.2.6 #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Neural and Evolutionary Computing (cs.NE) #Reinforcement Learning in Robotics #cs.AI #cs.LG #cs.NE #stat.ML
paper · pdf · doi:10.48550/arxiv.1910.04098
Accepted to ICLR 2020
openalex publication_date 2019/10/09 · arxiv created 2020/02/14 · arxiv updated 2020/02/17 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28
Biological evolution has distilled the experiences of many learners into the general learning algorithms of humans. Our novel meta reinforcement learning algorithm MetaGenRL is inspired by this process. MetaGenRL distills the experiences of many complex agents to meta-learn a low-complexity neural objective function that decides how future individuals will learn. Unlike recent meta-RL algorithms, MetaGenRL can generalize to new environments that are entirely different from those used for meta-training. In some cases, it even outperforms human-engineered RL algorithms. MetaGenRL uses off-policy second-order gradients during meta-training that greatly increase its sample efficiency.