Beutel, Alex
- The Case for Learned Index Structures
2017/12/04 by Tim Kraska, Alex Beutel, Kraska, Tim +7 · 10 voices · 23 citations
#cs.DB #cs.DS #cs.NE
- Underspecification Presents Challenges for Credibility in Modern Machine Learning
2020/11/06 by Alexander D'Amour, Alexander D’Amour, Katherine Heller +81 · 8 voices · 42 citations
Computer Science · Mathematics · #Explainable Artificial Intelligence (XAI) #Machine Learning in Healthcare #Topic Modeling #cs.LG #stat.ML
- The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions
2024/04/19 by Eric Wallace, Kai Xiao, Wallace, Eric +9 · 10 voices · 74 citations
Computer Science · Social Sciences · #Artificial Intelligence in Law #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Legal Education and Practice Innovations #Machine Learning (cs.LG) #cs.CL #cs.CR #cs.LG
- GPT-4o System Card
2024/10/25 by OpenAI, A. M. Hurst, : +503 · 1173 citations
Medicine · #Cardiovascular Function and Risk Factors #Hyperglycemia and glycemic control in critically ill and hospitalized patients
- OpenAI o1 System Card
2024/12/21 by OpenAI, Aaron Jaech, : +347 · 493 citations
Computer Science · #Advanced Computational Techniques and Applications
- Deliberative Alignment: Reasoning Enables Safer Language Models
2024/12/20 by Guan, Melody Y., Joglekar, Manas, Wallace, Eric +12 · 74 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computers and Society (cs.CY) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- MoB Queen: Mixture of Bandits with a Global Density Map
2018/12/06 by Minmin Chen, Chen, Minmin, Alex Beutel +9 · 18 citations
Computer Science · Decision Sciences · Engineering · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Information Retrieval (cs.IR) #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Smart Grid Energy Management
- Measuring and Reducing Gendered Correlations in Pre-trained Models
2020/10/12 by Kellie Webster, Webster, Kellie, Xuezhi Wang +13 · 20 citations
Social Sciences · #Ethics and Social Impacts of AI #Artificial Intelligence in Law
- HealthBench: Evaluating Large Language Models Towards Improved Human Health
2025/05/13 by Arora, Rahul K., Wei, Jason, Hicks, Rebecca Soskin +9 · 90 citations
#Computation and Language (cs.CL) #FOS: Computer and information sciences
- Data Decisions and Theoretical Implications when Adversarially Learning Fair Representations
2017/07/01 by Alex Beutel, Jilin Chen, Beutel, Alex +4 · 17 citations
Computer Science · Social Sciences · #Adversarial Robustness in Machine Learning #Computers and Society (cs.CY) #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Fairness without Demographics through Adversarially Reweighted Learning
2020/06/23 by Preethi Lahoti, Alex Beutel, Lahoti, Preethi +13 · 13 citations
Computer Science · Social Sciences · #Adversarial Robustness in Machine Learning #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Counterfactual Fairness in Text Classification through Robustness
2018/09/27 by Garg, Sahaj, Perot, Vincent, Limtiaco, Nicole +3 · 9 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Controlled Decoding from Language Models
2023/10/25 by Mudgal, Sidharth, Lee, Jong, Ganapathy, Harish +10 · 17 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Rule Based Rewards for Language Model Safety
2024/11/02 by Tong Mu, Mu, Tong, Alec Helyar +17 · 28 citations
Computer Science · Engineering · Social Sciences · #Access Control and Trust #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Safety Systems Engineering in Autonomy #Software Reliability and Analysis Research
- Fairness in Recommendation Ranking through Pairwise Comparisons
2019/03/02 by Beutel, Alex, Chen, Jilin, Doshi, Tulsee +8 · 8 citations
#Artificial Intelligence (cs.AI) #Computers and Society (cs.CY) #FOS: Computer and information sciences #Information Retrieval (cs.IR) #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Improving Diversity of Demographic Representation in Large Language Models via Collective-Critiques and Self-Voting
2023/10/25 by Preethi Lahoti, Nicholas Blumm, Lahoti, Preethi +19 · 8 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Speech and dialogue systems #Topic Modeling
- Putting Fairness Principles into Practice: Challenges, Metrics, and Improvements
2019/01/14 by Beutel, Alex, Chen, Jilin, Doshi, Tulsee +6 · 4 citations
#Artificial Intelligence (cs.AI) #Computers and Society (cs.CY) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Toward a better trade-off between performance and fairness with kernel-based distribution matching
2019/10/25 by Prost, Flavien, Qian, Hai, Chen, Qiuwen +3 · 3 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Measuring Recommender System Effects with Simulated Users
2021/01/12 by Sirui Yao, Yoni Halpern, Yao, Sirui +15 · 3 citations
Computer Science · Decision Sciences · Business, Management and Accounting · #Recommender Systems and Techniques #Advanced Bandit Algorithms Research #Consumer Market Behavior and Pricing
- Evaluating Fairness of Machine Learning Models Under Uncertain and Incomplete Information
2021/02/16 by Pranjal Awasthi, Alex Beutel, Awasthi, Pranjal +7 · 3 citations
Computer Science · Social Sciences · #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Privacy-Preserving Technologies in Data
- EdgeCentric: Anomaly Detection in Edge-Attributed Networks
2015/10/19 by Shah, Neil, Beutel, Alex, Hooi, Bryan +5 · 2 citations
#FOS: Computer and information sciences #Information Retrieval (cs.IR) #Social and Information Networks (cs.SI)
- From Hard Refusals to Safe-Completions: Toward Output-Centric Safety Training
2025/08/12 by Yuan, Yuan, Sriskandarajah, Tina, Brakman, Anna-Luisa +4 · 17 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computers and Society (cs.CY) #FOS: Computer and information sciences
- Diverse and Effective Red Teaming with Auto-generated Rewards and Multi-step Reinforcement Learning
2024/12/24 by Beutel, Alex, Xiao, Kai, Heidecke, Johannes +1 · 6 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- First-Person Fairness in Chatbots
2024/10/16 by Tyna Eloundou, Alex Beutel, Eloundou, Tyna +17 · 5 citations
Social Sciences · #Ethics and Social Impacts of AI #Digital Economy and Work Transformation
- Spotting Suspicious Link Behavior with fBox: An Adversarial Perspective
2014/10/15 by Neil Shah, Shah, Neil, Alex Beutel +5 · 1 citation
Computer Science · #Advanced Malware Detection Techniques #FOS: Computer and information sciences #Information Retrieval (cs.IR) #Machine Learning (cs.LG) #Network Security and Intrusion Detection #Social and Information Networks (cs.SI) #Spam and Phishing Detection
- BIRDNEST: Bayesian Inference for Ratings-Fraud Detection
2015/11/19 by Hooi, Bryan, Shah, Neil, Beutel, Alex +5 · 1 citation
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Social and Information Networks (cs.SI)
- Transfer of Machine Learning Fairness across Domains
2019/06/24 by Schumann, Candice, Wang, Xuezhi, Beutel, Alex +3 · 1 citation
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Improving Calibration through the Relationship with Adversarial Robustness
2020/06/29 by Qin, Yao, Wang, Xuezhi, Beutel, Alex +1 · 1 citation
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- CAT-Gen: Improving Robustness in NLP Models via Controlled Adversarial Text Generation
2020/10/05 by Tianlu Wang, Wang, Tianlu, Xuezhi Wang +13 · 1 citation
Computer Science · #Adversarial Robustness in Machine Learning #Computation and Language (cs.CL) #FOS: Computer and information sciences #Hate Speech and Cyberbullying Detection #Topic Modeling
- Understanding and Improving Robustness of Vision Transformers through Patch-based Negative Augmentation
2021/10/15 by Qin, Yao, Zhang, Chiyuan, Chen, Ting +3 · 1 citation
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- What Are Effective Labels for Augmented Data? Improving Calibration and Robustness with AutoLabel
2023/02/22 by Yao Qin, Xuezhi Wang, Qin, Yao +7 · 1 citation
Computer Science · #Advanced Neural Network Applications #FOS: Computer and information sciences #Image Enhancement Techniques #Machine Learning (cs.LG) #Machine Learning and Data Classification
- Striving for data-model efficiency: Identifying data externalities on group performance
2022/11/11 by Esther Rolf, Ben Packer, Rolf, Esther +5 · 1 citation
Computer Science · Decision Sciences · #Adversarial Robustness in Machine Learning #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Forecasting Techniques and Applications #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Effective Robustness against Natural Distribution Shifts for Models with Different Training Data
2023/02/02 by Zhouxing Shi, Shi, Zhouxing, Nicholas Carlini +11 · 1 citation
Computer Science · #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimodal Machine Learning Applications
- Let's Do a Thought Experiment: Using Counterfactuals to Improve Moral Reasoning
2023/06/25 by Xiao Ma, Swaroop Mishra, Ma, Xiao +7 · 1 citation
Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Topic Modeling
- Improving Few-shot Generalization of Safety Classifiers via Data Augmented Parameter-Efficient Fine-Tuning
2023/10/25 by Ananth Balashankar, Xiao Ma, Balashankar, Ananth +11 · 1 citation
Computer Science · Decision Sciences · Health Professions · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Occupational Health and Safety Research #Risk and Safety Analysis #Software Reliability and Analysis Research