Sean Wu
- Measuring Epistemic Resilience of LLMs Under Misleading Medical Context
2026/06/10 by Hongjian Zhou, Xinyu Zou, Jinge Wu +19 · 1 voice · 1 citation
Computer Science · #cs.CL
- Scientific reasoning does not reliably translate into scientific forecasting in frontier AI
2026/05/21 by Sean Wu, Pan Lu, Yupeng Chen +7 · 1 voice
Computer Science · #cs.AI
- BAS: A Decision-Theoretic Approach to Evaluating Large Language Model Confidence
2026/04/03 by Sean Wu, Fredrik K. Gustafsson, Edward Phillips +3 · 1 voice · 1 citation
Computer Science · #cs.CL
- BioMedArena: An Open-source Toolkit for Building and Evaluating Biomedical Deep Research Agents
2026/05/07 by Jinge Wu, Hongjian Zhou, Mingde Zeng +8 · 1 voice
Computer Science · #cs.AI