vix.ing · top · new · best · stats · spec

da Silva, Bruno Castro

  1. RLHF Deciphered: A Critical Analysis of Reinforcement Learning from Human Feedback for LLMs
    2024/04/12 by Shreyas Chaudhari, Pranjal Aggarwal, Chaudhari, Shreyas +13 · 26 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Topic Modeling
  2. Behavior Alignment via Reward Function Optimization
    2023/10/29 by Gupta, Dhawal, Chandak, Yash, Jordan, Scott M. +2 · 5 citations
    #FOS: Computer and information sciences #Machine Learning (cs.LG)
  3. On Ensuring that Intelligent Machines Are Well-Behaved
    2017/08/17 by Philip S. Thomas, Bruno Castro da Silva, Thomas, Philip S. +5 · 2 citations
    Computer Science · #Reinforcement Learning in Robotics #Explainable Artificial Intelligence (XAI) #Machine Learning and Data Classification
  4. Position: Benchmarking is Limited in Reinforcement Learning Research
    2024/06/23 by Scott M. Jordan, Adam White, Jordan, Scott M. +7 · 3 citations
    Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Methodology (stat.ME) #Open Source Software Innovations
  5. Enforcing Delayed-Impact Fairness Guarantees
    2022/08/24 by Weber, Aline, Metevier, Blossom, Brun, Yuriy +2 · 2 citations
    #Artificial Intelligence (cs.AI) #Computers and Society (cs.CY) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  6. Universal Off-Policy Evaluation
    2021/04/26 by Yash Chandak, Scott Niekum, Chandak, Yash +9 · 2 citations
    Decision Sciences · Mathematics · #Risk and Portfolio Optimization #Statistical Methods and Inference #Advanced Bandit Algorithms Research
  7. From Past to Future: Rethinking Eligibility Traces
    2023/12/20 by Gupta, Dhawal, Jordan, Scott M., Chaudhari, Shreyas +3 · 1 citation
    #FOS: Computer and information sciences #Machine Learning (cs.LG)