da Silva, Bruno Castro
- RLHF Deciphered: A Critical Analysis of Reinforcement Learning from Human Feedback for LLMs
2024/04/12 by Shreyas Chaudhari, Pranjal Aggarwal, Chaudhari, Shreyas +13 · 26 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Topic Modeling
- Behavior Alignment via Reward Function Optimization
2023/10/29 by Gupta, Dhawal, Chandak, Yash, Jordan, Scott M. +2 · 5 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG)
- On Ensuring that Intelligent Machines Are Well-Behaved
2017/08/17 by Philip S. Thomas, Bruno Castro da Silva, Thomas, Philip S. +5 · 2 citations
Computer Science · #Reinforcement Learning in Robotics #Explainable Artificial Intelligence (XAI) #Machine Learning and Data Classification
- Position: Benchmarking is Limited in Reinforcement Learning Research
2024/06/23 by Scott M. Jordan, Adam White, Jordan, Scott M. +7 · 3 citations
Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Methodology (stat.ME) #Open Source Software Innovations
- Enforcing Delayed-Impact Fairness Guarantees
2022/08/24 by Weber, Aline, Metevier, Blossom, Brun, Yuriy +2 · 2 citations
#Artificial Intelligence (cs.AI) #Computers and Society (cs.CY) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Universal Off-Policy Evaluation
2021/04/26 by Yash Chandak, Scott Niekum, Chandak, Yash +9 · 2 citations
Decision Sciences · Mathematics · #Risk and Portfolio Optimization #Statistical Methods and Inference #Advanced Bandit Algorithms Research
- From Past to Future: Rethinking Eligibility Traces
2023/12/20 by Gupta, Dhawal, Jordan, Scott M., Chaudhari, Shreyas +3 · 1 citation
#FOS: Computer and information sciences #Machine Learning (cs.LG)