vix.ing · top · new · best · stats · spec

Kasirzadeh, Atoosa

  1. Ethical and social risks of harm from Language Models
    2021/12/08 by Laura Weidinger, Weidinger, Laura, John W. Mellor +45 · 2 voices · 108 citations
    Computer Science · #cs.CL #cs.AI #cs.CY
  2. Multi-Agent Risks from Advanced AI
    2025/02/19 by Lewis Hammond, Alan Chan, Hammond, Lewis +89 · 3 voices · 38 citations
    Social Sciences · #Ethics and Social Impacts of AI #cs.AI #cs.CY #cs.ET #cs.LG #cs.MA
  3. Foundational Challenges in Assuring Alignment and Safety of Large Language Models
    2024/04/15 by Anwar, Usman, Saparov, Abulhair, Rando, Javier +39 · 23 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computers and Society (cs.CY) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  4. A Review of Modern Recommender Systems Using Generative Models (Gen-RecSys)
    2024/03/31 by Deldjoo, Yashar, He, Zhankui, McAuley, Julian +7 · 18 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Information Retrieval (cs.IR)
  5. Two Types of AI Existential Risk: Decisive and Accumulative
    2024/01/15 by Atoosa Kasirzadeh, Kasirzadeh, Atoosa · 3 voices · 8 citations
    Physics and Astronomy · Social Sciences · #Space Science and Extraterrestrial Life #Innovation, Sustainability, Human-Machine Systems
  6. Typology of Risks of Generative Text-to-Image Models
    2023/07/08 by Charlotte Bird, Bird, Charlotte, Eddie L. Ungless +3 · 8 citations
    Computer Science · Decision Sciences · #Computers and Society (cs.CY) #Data Visualization and Analytics #FOS: Computer and information sciences #Scientific Computing and Data Management
  7. In conversation with Artificial Intelligence: aligning language models with human values
    2022/09/01 by Kasirzadeh, Atoosa, Gabriel, Iason · 7 citations
    #Computation and Language (cs.CL) #Computers and Society (cs.CY) #FOS: Computer and information sciences
  8. Characterizing AI Agents for Alignment and Governance
    2025/04/30 by Atoosa Kasirzadeh, Kasirzadeh, Atoosa, Iason Gabriel +1 · 1 voice · 10 citations
    Computer Science · Engineering · Psychology · Social Sciences · #Artificial Intelligence (cs.AI) #Computers and Society (cs.CY) #Ethics and Social Impacts of AI #FOS: Computer and information sciences #FOS: Electrical engineering #Human-Automation Interaction and Safety #Multi-Agent Systems and Negotiation #Systems and Control (eess.SY) #cs.AI #cs.CY #eess.SY #electronic engineering #information engineering
  9. The Use and Misuse of Counterfactuals in Ethical Machine Learning
    2021/02/09 by Kasirzadeh, Atoosa, Smart, Andrew · 3 citations
    #Computers and Society (cs.CY) #FOS: Computer and information sciences
  10. Discipline and Label: A WEIRD Genealogy and Social Theory of Data Annotation
    2024/02/09 by Smart, Andrew, Wang, Ding, Monk, Ellis +4 · 4 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences
  11. A Taxonomy of Systemic Risks from General-Purpose AI
    2024/11/24 by Risto Uuk, Uuk, Risto, Carlos Ignacio Gutierrez +13 · 1 voice · 3 citations
    Decision Sciences · Engineering · #Fault Detection and Control Systems #Risk and Safety Analysis #cs.CY
  12. Epistemic Injustice in Generative AI
    2024/08/21 by Jackie Kay, Atoosa Kasirzadeh, Kay, Jackie +3 · 3 citations
    Arts and Humanities · Neuroscience · Social Sciences · #Artificial Intelligence (cs.AI) #Epistemology, Ethics, and Metaphysics #Ethics and Social Impacts of AI #FOS: Computer and information sciences #Psychology of Moral and Emotional Judgment
  13. The Future of Open Human Feedback
    2024/08/15 by Shachar Don-Yehiya, Don-Yehiya, Shachar, Ben Burtenshaw +37 · 4 citations
    Psychology · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Human-Automation Interaction and Safety #Human-Computer Interaction (cs.HC)
  14. Democratic AI is Possible. The Democracy Levels Framework Shows How It Might Work
    2024/11/14 by Aviv Ovadya, Ovadya, Aviv, Kyle Redman +19 · 3 citations
    Social Sciences · #Computers and Society (cs.CY) #Ethics and Social Impacts of AI #FOS: Computer and information sciences
  15. AI Safety for Everyone
    2025/02/13 by Bálint Gyevnár, Gyevnar, Balint, Atoosa Kasirzadeh +1 · 4 citations
    Social Sciences · #Computers and Society (cs.CY) #Ethics and Social Impacts of AI #FOS: Computer and information sciences
  16. AI, Digital Platforms, and the New Systemic Risk
    2025/09/22 by Philipp Hacker, Hacker, Philipp, Lilian Edwards +3 · 1 voice · 2 citations
    Social Sciences · Economics, Econometrics and Finance · #Economic and Technological Developments in Russia #Economic Development and Digital Transformation
  17. Generative Value Conflicts Reveal LLM Priorities
    2025/09/29 by Andy Liu, Kshitish Ghate, Liu, Andy +9 · 1 voice · 2 citations
    #cs.CL #cs.AI #cs.LG
  18. CIVICS: Building a Dataset for Examining Culturally-Informed Values in Large Language Models
    2024/05/22 by Pistilli, Giada, Leidinger, Alina, Jernite, Yacine +3 · 1 citation
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
  19. Measurement challenges in AI catastrophic risk governance and safety frameworks
    2024/10/01 by Atoosa Kasirzadeh, Kasirzadeh, Atoosa · 1 citation
    Business, Management and Accounting · Social Sciences · Health Professions · #Supply Chain Resilience and Risk Management #Ethics and Social Impacts of AI #Occupational Health and Safety Research
  20. The More You Automate, the Less You See: Hidden Pitfalls of AI Scientist Systems
    2025/09/10 by Luo, Ziming, Atoosa Kasirzadeh, Nihar B. Shah +2 · 2 citations
    Computer Science · Decision Sciences · Social Sciences · #Artificial Intelligence (cs.AI) #Digital Libraries (cs.DL) #Ethics and Social Impacts of AI #FOS: Computer and information sciences #Research Data Management Practices #Scientific Computing and Data Management
  21. Explanation Hacking: The perils of algorithmic recourse
    2024/03/22 by Emily Sullivan, Sullivan, Emily, Atoosa Kasirzadeh +1 · 1 citation
    Social Sciences · #Ethics and Social Impacts of AI
  22. Ethics Whitepaper: Whitepaper on Ethical Research into Large Language Models
    2024/10/17 by Ungless, Eddie L., Vitsakis, Nikolas, Talat, Zeerak +5 · 1 citation
    #Computation and Language (cs.CL) #Computers and Society (cs.CY) #FOS: Computer and information sciences #I.2
  23. The Only Way is Ethics: A Guide to Ethical Research with Large Language Models
    2024/12/20 by Eddie L. Ungless, Nikolas Vitsakis, Ungless, Eddie L. +13 · 1 citation
    Social Sciences · #Social Science and Policy Research