2023/09/05 by Max Tegmark, Steve Omohundro, Tegmark, Max +1 · 2 voices · 17 citations
Computer Science · Engineering · Mathematics · #Advanced Malware Detection Techniques #Adversarial Robustness in Machine Learning #Artificial intelligence #Computability, Logic, AI Algorithms #Computer science #Engineering #Humanity #Interpretability #Law #Mathematical economics #Mathematics #Outcome (game theory) #Path (computing) #Political science #Programming language #Sociology #Theoretical computer science #Thriving #Work (physics) #cs.AI #cs.CY #cs.LG
paper · pdf · doi:10.48550/arxiv.2309.01933
published in arXiv (Cornell University) (Cornell University)
openalex publication_date 2023/09/05 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28
We describe a path to humanity safely thriving with powerful Artificial General Intelligences (AGIs) by building them to provably satisfy human-specified requirements. We argue that this will soon be technically feasible using advanced AI for formal verification and mechanistic interpretability. We further argue that it is the only path which guarantees safe controlled AGI. We end with a list of challenge problems whose solution would contribute to this positive outcome and invite readers to join in this work.