vix.ing · top · new · best · stats · spec

Coping with the variability in humans reward during simulated\n human-robot interactions through the coordination of multiple learning\n strategies

2020/05/06 by Rémi Dromnelle, Dromnelle, Rémi, Benoît Girard +7 · 1 citation
Computer Science · Engineering · #FOS: Computer and information sciences #Reinforcement Learning in Robotics #Robot Manipulation and Learning #Robotic Locomotion and Control #Robotics (cs.RO)

paper · pdf · doi:10.48550/arxiv.2005.03987

openalex publication_date 2020/05/06 · openalex created_date 2022/07/26 · openalex updated_date 2026/07/28

Abstract

An important current challenge in Human-Robot Interaction (HRI) is to enable\nrobots to learn on-the-fly from human feedback. However, humans show a great\nvariability in the way they reward robots. We propose to address this issue by\nenabling the robot to combine different learning strategies, namely model-based\n(MB) and model-free (MF) reinforcement learning. We simulate two HRI scenarios:\na simple task where the human congratulates the robot for putting the right\ncubes in the right boxes, and a more complicated version of this task where\ncubes have to be placed in a specific order. We show that our existing MB-MF\ncoordination algorithm previously tested in robot navigation works well here\nwithout retuning parameters. It leads to the maximal performance while\nproducing the same minimal computational cost as MF alone. Moreover, the\nalgorithm gives a robust performance no matter the variability of the simulated\nhuman feedback, while each strategy alone is impacted by this variability.\nOverall, the results suggest a promising way to promote robot learning\nflexibility when facing variable human feedback.\n

Cited by

Related