1995/01/01 by Andrew G. Barto, Steven J. Bradtke, Satinder Singh +1 · 1,113 citations
Computer Science · #AI-based Problem Solving and Planning #Algorithm #Artificial intelligence #Asynchronous communication #Bridge (graph theory) #Computer science #Control (management) #Dynamic programming #Evolutionary Algorithms and Applications #Machine learning #Reinforcement Learning in Robotics #Reinforcement learning #Temporal difference learning
paper · pdf · doi:10.1016/0004-3702(94)00011-o
published in Artificial Intelligence 72(1-2), 81-138 (Elsevier BV)
openalex publication_date 1995/01/01 · openalex created_date 2025/10/10 · openalex updated_date 2026/08/03