Lawrence Chan
- Measuring AI Ability to Complete Long Software Tasks
2025/03/18 by Thomas Kwa, Ben West, Kwa, Thomas +48 · 26 voices · 40 citations
#cs.AI #cs.LG
- Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
2025/07/07 by Gheorghe Comanici, Comanici, Gheorghe, Eric Bieber +6844 · 8 voices · 1374 citations
#cs.CL #cs.AI
- Progress measures for grokking via mechanistic interpretability
2023/01/12 by Neel Nanda, Nanda, Neel, Lawrence Chan +8 · 3 voices · 106 citations
Computer Science · Engineering · Neuroscience · #Advanced Memory and Neural Computing #Neural Networks and Applications #Neural dynamics and brain function #cs.AI #cs.LG
- RE-Bench: Evaluating frontier AI R&D capabilities of language model agents against human experts
2024/11/22 by Hjalmar Wijk, Wijk, Hjalmar, Tao Lin +47 · 3 voices · 13 citations
Computer Science · Medicine · Social Sciences · #Artificial Intelligence in Healthcare and Education #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #cs.AI #cs.LG
- Mathematical Models of Computation in Superposition
2024/08/10 by Kaarel Hänni, Jake Mendel, Hänni, Kaarel +5 · 7 citations
Computer Science · #Computability, Logic, AI Algorithms
- Adversarial Training for High-Stakes Reliability
2022/05/03 by Daniel M. Ziegler, Seraphina Nix, Ziegler, Daniel M. +21 · 3 citations
Computer Science · Social Sciences · #Adversarial Robustness in Machine Learning #Explainable Artificial Intelligence (XAI) #Ethics and Social Impacts of AI
- Language models are better than humans at next-token prediction
2022/12/21 by Buck Shlegeris, Shlegeris, Buck, Fabien Roger +5 · 3 citations
Computer Science · #Topic Modeling #Text Readability and Simplification #Natural Language Processing Techniques
- HCAST: Human-Calibrated Autonomy Software Tasks
2025/03/21 by David B. Rein, Rein, David, Becker, Joel +38 · 3 citations
Social Sciences · Computer Science · #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #Adversarial Robustness in Machine Learning
- Targeted inactivation of MLL3 histone H3–Lys-4 methyltransferase activity in the mouse reveals vital roles for MLL3 in adipogenesis
2008/12/01 by Jeongkyung Lee, Pradip Saha, Pradip K. Saha +9 · 25 citations
Biochemistry, Genetics and Molecular Biology · Immunology and Microbiology · #Epigenetics and DNA Methylation #Immune Cell Function and Interaction #RNA modifications and cancer
- Compact Proofs of Model Performance via Mechanistic Interpretability
2024/06/17 by Jason N. Gross, Rajashree Agrawal, Gross, Jason +13 · 1 citation
Computer Science · Physics and Astronomy · #FOS: Computer and information sciences #Logic in Computer Science (cs.LO) #Machine Learning (cs.LG) #Machine Learning and Algorithms #Model Reduction and Neural Networks