vix.ing · top · new · best · stats · spec

Andrew M. Bean

  1. Clinical knowledge in LLMs does not translate to human interactions
    2025/04/26 by Andrew M. Bean, Rebecca Payne, Bean, Andrew M. +19 · 15 voices · 5 citations
    Medicine · Computer Science · #Artificial Intelligence in Healthcare and Education #Global Health and Surgery #Explainable Artificial Intelligence (XAI)
  2. LINGOLY: A Benchmark of Olympiad-Level Linguistic Reasoning Puzzles in Low-Resource and Extinct Languages
    2024/06/10 by Andrew M. Bean, Simi Hellsten, Bean, Andrew M. +13 · 9 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques
  3. The Past, Present and Better Future of Feedback Learning in Large Language Models for Subjective Human Preferences and Values
    2023/10/11 by Hannah Rose Kirk, Andrew M. Bean, Kirk, Hannah Rose +7 · 5 citations
    Computer Science · #Topic Modeling #Natural Language Processing Techniques #Explainable Artificial Intelligence (XAI)
  4. Measuring what Matters: Construct Validity in Large Language Model Benchmarks
    2025/11/03 by Andrew M. Bean, Bean, Andrew M., Angelika Romanou +78 · 10 citations
    Computer Science · Medicine · Social Sciences · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Healthcare and Education #Computation and Language (cs.CL) #Computational and Text Analysis Methods #FOS: Computer and information sciences #Topic Modeling
  5. Reliability of LLMs as medical assistants for the general public: a randomized preregistered study
    2026/02/01 by Andrew M. Bean, Rebecca Payne, Guy Parsons +8 · 2 voices · 14 citations
    Medicine · Computer Science · #Artificial Intelligence in Healthcare and Education #Global Health and Surgery #Persona Design and Applications