vix.ing · top · new · best · stats · spec

Evaluating AI Providers' Frontier Safety Frameworks

2025/12/01 by Lily Stelling, Stelling, Lily, Melanie Murray +7 · 1 citation
Health Professions · Medicine · Social Sciences · #Artificial Intelligence in Healthcare and Education #Computers and Society (cs.CY) #Ethics and Social Impacts of AI #FOS: Computer and information sciences #Occupational Health and Safety Research

paper · pdf · doi:10.48550/arxiv.2512.01166

openalex publication_date 2025/12/01 · openalex created_date 2025/12/03 · openalex updated_date 2026/07/31

Abstract

Following the AI Seoul Summit in 2024, twelve AI companies published frontier AI safety frameworks (Frameworks) outlining their approaches to managing catastrophic risks from advanced AI systems. Emerging legislation increasingly treats these Frameworks as external accountability mechanisms, incorporating them into reporting requirements. But what do the Frameworks actually commit each company to do? This study assesses 12 Frameworks, using 65 weighted criteria, across four dimensions: risk identification, risk analysis & evaluation, risk treatment, and risk governance. Our criteria adapt established risk management principles from other high-risk industries (e.g. aviation, nuclear power) to the frontier AI context, following Campos et al. (2025). Overall scores range from 34% (Anthropic) to 8% (Cohere), with a median of 18%. Many aspects are missing or under-specified. These low scores may be natural given the nascency of AI risk management compared to industries with decades of practice. Nonetheless, current Frameworks are limited as accountability functions, with vague commitments that make it difficult to predict company decisions, assess whether planned responses are adequate, or determine whether commitments have been kept. Still, higher scores appear feasible within current constraints: a company adopting all leading practices currently adopted across their peers would score 54%, which is triple the current median.

Cited by

Related