Bean, Andrew M.
- Clinical knowledge in LLMs does not translate to human interactions
2025/04/26 by Andrew M. Bean, Bean, Andrew M., Rebecca Payne +19 · 15 voices · 5 citations
Medicine · Computer Science · #Artificial Intelligence in Healthcare and Education #Global Health and Surgery #Explainable Artificial Intelligence (XAI)
- LINGOLY: A Benchmark of Olympiad-Level Linguistic Reasoning Puzzles in Low-Resource and Extinct Languages
2024/06/10 by Andrew M. Bean, Bean, Andrew M., Simi Hellsten +13 · 9 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques
- The Past, Present and Better Future of Feedback Learning in Large Language Models for Subjective Human Preferences and Values
2023/10/11 by Hannah Rose Kirk, Kirk, Hannah Rose, Andrew M. Bean +7 · 5 citations
Computer Science · #Topic Modeling #Natural Language Processing Techniques #Explainable Artificial Intelligence (XAI)
- Measuring what Matters: Construct Validity in Large Language Model Benchmarks
2025/11/03 by Andrew M. Bean, Bean, Andrew M., Kearns, Ryan Othniel +78 · 10 citations
Computer Science · Medicine · Social Sciences · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Healthcare and Education #Computation and Language (cs.CL) #Computational and Text Analysis Methods #FOS: Computer and information sciences #Topic Modeling
- LLMs Don't Know Their Own Decision Boundaries: The Unreliability of Self-Generated Counterfactual Explanations
2025/09/11 by Mayne, Harry, Kearns, Ryan Othniel, Yang, Yushi +4 · 1 voice · 2 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Do Large Language Models have Shared Weaknesses in Medical Question Answering?
2023/10/11 by Bean, Andrew M., Korgul, Karolina, Krones, Felix +2 · 1 citation
#Computation and Language (cs.CL) #FOS: Computer and information sciences