Sayash Kapoor
- The Leaderboard Illusion
2025/04/29 by Shivalika Singh, Yiyang Nan, Singh, Shivalika +26 · 32 voices · 19 citations
Medicine · Computer Science · Social Sciences · #Artificial Intelligence in Healthcare and Education #AI in Service Interactions #Ethics and Social Impacts of AI
- AI Agents That Matter
2024/07/01 by Sayash Kapoor, Kapoor, Sayash, Benedikt Stroebl +7 · 4 voices · 29 citations
#cs.LG #cs.AI
- The Limits of Inference Scaling Through Resampling
2024/11/26 by Benedikt Stroebl, Sayash Kapoor, Stroebl, Benedikt +3 · 2 voices · 16 citations
Computer Science · #Natural Language Processing Techniques #Topic Modeling
- Holistic Agent Leaderboard: The Missing Infrastructure for AI Agent Evaluation
2025/10/13 by Sayash Kapoor, Benedikt Stroebl, Kapoor, Sayash +63 · 3 voices · 9 citations
Computer Science · #Multi-Agent Systems and Negotiation
- A Safe Harbor for AI Evaluation and Red Teaming
2024/03/07 by Shayne Longpre, Longpre, Shayne, Sayash Kapoor +43 · 1 voice · 15 citations
Computer Science · #Explainable Artificial Intelligence (XAI)
- On the Societal Impact of Open Foundation Models
2024/02/27 by Sayash Kapoor, Rishi Bommasani, Kapoor, Sayash +47 · 15 citations
Engineering · #3D Modeling in Geospatial Applications #Artificial Intelligence (cs.AI) #BIM and Construction Integration #Computers and Society (cs.CY) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- International AI Safety Report
2025/01/29 by Yoshua Bengio, Bengio, Yoshua, Sören Mindermann +180 · 17 citations
Social Sciences · #Artificial Intelligence (cs.AI) #Computers and Society (cs.CY) #Ethics and Social Impacts of AI #FOS: Computer and information sciences #Machine Learning (cs.LG)
- The Foundation Model Transparency Index
2023/10/19 by Rishi Bommasani, Bommasani, Rishi, Kevin Klyman +13 · 1 voice · 8 citations
Computer Science · #Blockchain Technology Applications and Security #Mobile Crowdsensing and Crowdsourcing #Privacy-Preserving Technologies in Data #cs.AI #cs.LG
- Towards a Science of AI Agent Reliability
2026/02/18 by Stephan Rabanser, Sayash Kapoor, Peter Kirgis +3 · 6 voices · 2 citations
#cs.AI #cs.CY #cs.LG
- CORE-Bench: Fostering the Credibility of Published Research Through a Computational Reproducibility Agent Benchmark
2024/09/17 by Zachary S. Siegel, Siegel, Zachary S., Sayash Kapoor +7 · 8 citations
Decision Sciences · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Multiagent Systems (cs.MA) #Scientific Computing and Data Management
- Leakage and the Reproducibility Crisis in ML-based Science
2022/07/14 by Sayash Kapoor, Arvind Narayanan, Kapoor, Sayash +1 · 4 citations
Computer Science · #Anomaly Detection Techniques and Applications #Artificial Intelligence (cs.AI) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Data Classification #Methodology (stat.ME)
- Promises and pitfalls of artificial intelligence for legal applications
2024/01/10 by Sayash Kapoor, Kapoor, Sayash, Peter Henderson +3 · 4 citations
Computer Science · Social Sciences · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Law #Computers and Society (cs.CY) #FOS: Computer and information sciences #Law, AI, and Intellectual Property
- Can AI agents conduct open-ended AI research? Early evidence from two case studies
2026/07/29 by Peter Kirgis, Sayash Kapoor, Andrew Schwartz +21 · 2 voices
Computer Science · #cs.AI #cs.CY #cs.LG
- An Algorithmic Framework to Control Bias in Bandit-based Personalization
2018/02/23 by L. Elisa Celis, Celis, L. Elisa, Sayash Kapoor +5 · 1 citation
Decision Sciences · Engineering · Computer Science · #Advanced Bandit Algorithms Research #Smart Grid Energy Management #Recommender Systems and Techniques
- The Reality of AI and Biorisk
2024/12/02 by Aidan Peppin, Anka Reuel, Peppin, Aidan +21 · 1 citation
Computer Science · Engineering · Social Sciences · #Artificial Intelligence (cs.AI) #Digital Transformation in Industry #Ethics and Social Impacts of AI #FOS: Computer and information sciences #Physical Unclonable Functions (PUFs) and Hardware Security
- The 2025 Foundation Model Transparency Index
2025/12/11 by Alexander Wan, Kevin Klyman, Wan, Alexander +13 · 1 voice · 3 citations
Business, Management and Accounting · Computer Science · Social Sciences · #Big Data and Business Intelligence #Data Analysis with R #Ethics and Social Impacts of AI #cs.AI #cs.CY #cs.LG