vix.ing · top · new · best · stats · spec

Thomas Krendl Gilbert

  1. Open Problems and Fundamental Limitations of Reinforcement Learning from Human Feedback
    2023/07/27 by Stephen Casper, Casper, Stephen, Xander Davies +65 · 3 voices · 161 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Software Reliability and Analysis Research #cs.AI #cs.CL #cs.LG
  2. Toward Trustworthy AI Development: Mechanisms for Supporting Verifiable Claims
    2020/04/15 by Miles Brundage, Brundage, Miles, Shahar Avin +123 · 2 voices · 35 citations
    Computer Science · Social Sciences · #Adversarial Robustness in Machine Learning #Ethics and Social Impacts of AI #Law, AI, and Intellectual Property #cs.CY
  3. Beyond Bias and Compliance: Towards Individual Agency and Plurality of Ethics in AI
    2023/02/23 by Thomas Krendl Gilbert, Gilbert, Thomas Krendl, Megan Welle Brozek +3 · 1 voice · 1 citation
    Computer Science · Neuroscience · Social Sciences · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Computers and Society (cs.CY) #Ethics and Social Impacts of AI #FOS: Computer and information sciences #Psychology of Moral and Emotional Judgment #cs.AI #cs.CY
  4. A Broader View on Bias in Automated Decision-Making: Reflecting on Epistemology and Dynamics
    2018/07/02 by Roel Dobbe, Dobbe, Roel, Sarah Dean +5 · 2 citations
    Computer Science · Social Sciences · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Dynamical Systems (math.DS) #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #FOS: Electrical engineering #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Systems and Control (eess.SY) #electronic engineering #information engineering