vix.ing · top · new · best · stats · spec

Garfinkel, Ben

  1. On the Impossibility of Supersized Machines
    2017/03/31 by Ben Garfinkel, Garfinkel, Ben, Miles Brundage +15 · 11 voices
    #cs.CY #physics.pop-ph
  2. The Malicious Use of Artificial Intelligence: Forecasting, Prevention, and Mitigation
    2018/02/20 by Miles Brundage, Brundage, Miles, Shahar Avin +49 · 6 voices · 19 citations
    #cs.AI #cs.CR #cs.CY
  3. Model evaluation for extreme risks
    2023/05/24 by Toby Shevlane, Sebastian Farquhar, Shevlane, Toby +39 · 14 citations
    Computer Science · #Software Engineering Research #Software Reliability and Analysis Research #Information and Cyber Security
  4. International AI Safety Report
    2025/01/29 by Yoshua Bengio, Bengio, Yoshua, Sören Mindermann +180 · 17 citations
    Social Sciences · #Artificial Intelligence (cs.AI) #Computers and Society (cs.CY) #Ethics and Social Impacts of AI #FOS: Computer and information sciences #Machine Learning (cs.LG)
  5. Open-Sourcing Highly Capable Foundation Models: An evaluation of risks, benefits, and alternative methods for pursuing open-source objectives
    2023/09/29 by Elizabeth Seger, Noemi Dreksler, Seger, Elizabeth +41 · 7 citations
    Computer Science · Decision Sciences · Social Sciences · #Artificial Intelligence (cs.AI) #Computers and Society (cs.CY) #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Scientific Computing and Data Management #Software Engineering (cs.SE)
  6. Democratising AI: Multiple Meanings, Goals, and Methods
    2023/03/22 by Elizabeth Seger, Seger, Elizabeth, Aviv Ovadya +7 · 6 citations
    Social Sciences · #Ethics and Social Impacts of AI
  7. The Windfall Clause: Distributing the Benefits of AI for the Common Good
    2019/12/25 by Cullen O’Keefe, O'Keefe, Cullen, Peter Cihon +9 · 2 citations
    Social Sciences · Neuroscience · #Ethics and Social Impacts of AI #Neuroethics, Human Enhancement, Biomedical Innovations #Innovation, Sustainability, Human-Machine Systems
  8. Towards best practices in AGI safety and governance: A survey of expert opinion
    2023/05/11 by Jonas Schuett, Noemi Dreksler, Schuett, Jonas +11 · 3 citations
    Computer Science · Medicine · Social Sciences · #Artificial Intelligence in Healthcare and Education #Computers and Society (cs.CY) #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences
  9. From Principles to Rules: A Regulatory Approach for Frontier AI
    2024/07/10 by Schuett, Jonas, Anderljung, Markus, Carlier, Alexis +2 · 2 citations
    #Computers and Society (cs.CY) #FOS: Computer and information sciences
  10. Third-party compliance reviews for frontier AI safety frameworks
    2025/05/03 by Aidan Homewood, Sophie Williams, Homewood, Aidan +25 · 1 citation
    Computer Science · #Adversarial Robustness in Machine Learning #Computers and Society (cs.CY) #FOS: Computer and information sciences