Thomas Krendl Gilbert
- Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback
2023/07/27 by Stephen Casper, Casper, Stephen, Xander Davies +65 · 3 voices · 161 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Software Reliability and Analysis Research #cs.AI #cs.CL #cs.LG
- Toward Trustworthy AI Development: Mechanisms for Supporting Verifiable Claims
2020/04/15 by Miles Brundage, Brundage, Miles, Shahar Avin +123 · 2 voices · 35 citations
Computer Science · Social Sciences · #Adversarial Robustness in Machine Learning #Ethics and Social Impacts of AI #Law, AI, and Intellectual Property #cs.CY
- Beyond Bias and Compliance: Towards Individual Agency and Plurality of Ethics in AI
2023/02/23 by Thomas Krendl Gilbert, Gilbert, Thomas Krendl, Megan Welle Brozek +3 · 1 voice · 1 citation
Computer Science · Neuroscience · Social Sciences · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Computers and Society (cs.CY) #Ethics and Social Impacts of AI #FOS: Computer and information sciences #Psychology of Moral and Emotional Judgment #cs.AI #cs.CY
- A Broader View on Bias in Automated Decision-Making: Reflecting on Epistemology and Dynamics
2018/07/02 by Roel Dobbe, Dobbe, Roel, Sarah Dean +5 · 2 citations
Computer Science · Social Sciences · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Dynamical Systems (math.DS) #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #FOS: Electrical engineering #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Systems and Control (eess.SY) #electronic engineering #information engineering