Weiss, Rebecca
- Establishing Best Practices for Building Rigorous Agentic Benchmarks
2025/07/03 by Zhu, Yuxuan, Jin, Tengjun, Pruksachatkun, Yada +22 · 5 voices · 14 citations
#A.1 #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #I.2.m
- "This Browser is Lightning Fast": The Effects of Message Content on Perceived Performance
2021/03/10 by Jess Hohenstein, Hohenstein, Jess, Bill Selman +7 · 2 voices
#cs.HC
- Introducing v0.5 of the AI Safety Benchmark from MLCommons
2024/04/18 by Bertie Vidgen, Adarsh Agrawal, Vidgen, Bertie +202 · 2 voices · 15 citations
Computer Science · #Adversarial Robustness in Machine Learning
- AILuminate: Introducing v1.0 of the AI Risk and Reliability Benchmark from MLCommons
2025/02/19 by Shaona Ghosh, Heather Frase, Ghosh, Shaona +200 · 1 voice · 9 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computers and Society (cs.CY) #FOS: Computer and information sciences #cs.AI #cs.CY
- In-House Evaluation Is Not Enough: Towards Robust Third-Party Flaw Disclosure for General-Purpose AI
2025/03/21 by Longpre, Shayne, Klyman, Kevin, Appel, Ruth E. +31 · 6 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences
- Political Polarization in Online News Consumption
2021/04/09 by Garimella, Kiran, Smith, Tim, Weiss, Rebecca +1 · 1 citation
#Computers and Society (cs.CY) #FOS: Computer and information sciences #Social and Information Networks (cs.SI)
- Towards Best Practices for Open Datasets for LLM Training
2025/01/14 by Stefan Baack, Baack, Stefan, Stella Biderman +79 · 3 voices · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computers and Society (cs.CY) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #cs.AI #cs.CL #cs.CY #cs.LG