vix.ing · top · new · best · stats · spec

Booth, Serena

  1. Models of human preference for learning reward functions
    2022/06/05 by W. Bradley Knox, Stephane Hatgis-Kessell, Knox, W. Bradley +9 · 11 citations
    Psychology · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #FOS: Electrical engineering #I.2.6 #I.2.8 #Machine Learning (cs.LG) #Mental Health Research Topics #Systems and Control (eess.SY) #electronic engineering #information engineering
  2. Do Feature Attribution Methods Correctly Attribute Features?
    2021/04/27 by Yilun Zhou, Serena Booth, Zhou, Yilun +5 · 7 citations
    Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  3. Learning Optimal Advantage from Preferences and Mistaking it for Reward
    2023/10/03 by Knox, W. Bradley, Hatgis-Kessell, Stephane, Adalgeirsson, Sigurdur Orn +4 · 2 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #I.2.6 #I.2.8 #Machine Learning (cs.LG)
  4. Towards Improving Reward Design in RL: A Reward Alignment Metric for RL Practitioners
    2025/03/08 by Muslimani, Calarina, Johnstonbaugh, Kerrick, Chandramouli, Suyog +3 · 2 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)