vix.ing · top · new · best · stats · spec

Lawrence Chan

  1. Measuring AI Ability to Complete Long Software Tasks
    2025/03/18 by Thomas Kwa, Ben West, Kwa, Thomas +48 · 26 voices · 40 citations
    #cs.AI #cs.LG
  2. Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
    2025/07/07 by Gheorghe Comanici, Comanici, Gheorghe, Eric Bieber +6844 · 8 voices · 1374 citations
    #cs.CL #cs.AI
  3. Progress measures for grokking via mechanistic interpretability
    2023/01/12 by Neel Nanda, Nanda, Neel, Lawrence Chan +8 · 3 voices · 106 citations
    Computer Science · Engineering · Neuroscience · #Advanced Memory and Neural Computing #Neural Networks and Applications #Neural dynamics and brain function #cs.AI #cs.LG
  4. RE-Bench: Evaluating frontier AI R&D capabilities of language model agents against human experts
    2024/11/22 by Hjalmar Wijk, Wijk, Hjalmar, Tao Lin +47 · 3 voices · 13 citations
    Computer Science · Medicine · Social Sciences · #Artificial Intelligence in Healthcare and Education #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #cs.AI #cs.LG
  5. Mathematical Models of Computation in Superposition
    2024/08/10 by Kaarel Hänni, Jake Mendel, Hänni, Kaarel +5 · 7 citations
    Computer Science · #Computability, Logic, AI Algorithms
  6. Adversarial Training for High-Stakes Reliability
    2022/05/03 by Daniel M. Ziegler, Seraphina Nix, Ziegler, Daniel M. +21 · 3 citations
    Computer Science · Social Sciences · #Adversarial Robustness in Machine Learning #Explainable Artificial Intelligence (XAI) #Ethics and Social Impacts of AI
  7. Language models are better than humans at next-token prediction
    2022/12/21 by Buck Shlegeris, Shlegeris, Buck, Fabien Roger +5 · 3 citations
    Computer Science · #Topic Modeling #Text Readability and Simplification #Natural Language Processing Techniques
  8. HCAST: Human-Calibrated Autonomy Software Tasks
    2025/03/21 by David B. Rein, Rein, David, Becker, Joel +38 · 3 citations
    Social Sciences · Computer Science · #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #Adversarial Robustness in Machine Learning
  9. Targeted inactivation of MLL3 histone H3–Lys-4 methyltransferase activity in the mouse reveals vital roles for MLL3 in adipogenesis
    2008/12/01 by Jeongkyung Lee, Pradip Saha, Pradip K. Saha +9 · 25 citations
    Biochemistry, Genetics and Molecular Biology · Immunology and Microbiology · #Epigenetics and DNA Methylation #Immune Cell Function and Interaction #RNA modifications and cancer
  10. Compact Proofs of Model Performance via Mechanistic Interpretability
    2024/06/17 by Jason N. Gross, Rajashree Agrawal, Gross, Jason +13 · 1 citation
    Computer Science · Physics and Astronomy · #FOS: Computer and information sciences #Logic in Computer Science (cs.LO) #Machine Learning (cs.LG) #Machine Learning and Algorithms #Model Reduction and Neural Networks