vix.ing · top · new · best · stats · spec

Beutel, Alex

  1. The Case for Learned Index Structures
    2017/12/04 by Tim Kraska, Alex Beutel, Kraska, Tim +7 · 10 voices · 23 citations
    #cs.DB #cs.DS #cs.NE
  2. Underspecification Presents Challenges for Credibility in Modern Machine Learning
    2020/11/06 by Alexander D'Amour, Alexander D’Amour, Katherine Heller +81 · 8 voices · 42 citations
    Computer Science · Mathematics · #Explainable Artificial Intelligence (XAI) #Machine Learning in Healthcare #Topic Modeling #cs.LG #stat.ML
  3. The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions
    2024/04/19 by Eric Wallace, Kai Xiao, Wallace, Eric +9 · 10 voices · 74 citations
    Computer Science · Social Sciences · #Artificial Intelligence in Law #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Legal Education and Practice Innovations #Machine Learning (cs.LG) #cs.CL #cs.CR #cs.LG
  4. GPT-4o System Card
    2024/10/25 by OpenAI, A. M. Hurst, : +503 · 1173 citations
    Medicine · #Cardiovascular Function and Risk Factors #Hyperglycemia and glycemic control in critically ill and hospitalized patients
  5. OpenAI o1 System Card
    2024/12/21 by OpenAI, Aaron Jaech, : +347 · 493 citations
    Computer Science · #Advanced Computational Techniques and Applications
  6. Deliberative Alignment: Reasoning Enables Safer Language Models
    2024/12/20 by Guan, Melody Y., Joglekar, Manas, Wallace, Eric +12 · 74 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computers and Society (cs.CY) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  7. MoB Queen: Mixture of Bandits with a Global Density Map
    2018/12/06 by Minmin Chen, Chen, Minmin, Alex Beutel +9 · 18 citations
    Computer Science · Decision Sciences · Engineering · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Information Retrieval (cs.IR) #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Smart Grid Energy Management
  8. Measuring and Reducing Gendered Correlations in Pre-trained Models
    2020/10/12 by Kellie Webster, Webster, Kellie, Xuezhi Wang +13 · 20 citations
    Social Sciences · #Ethics and Social Impacts of AI #Artificial Intelligence in Law
  9. HealthBench: Evaluating Large Language Models Towards Improved Human Health
    2025/05/13 by Arora, Rahul K., Wei, Jason, Hicks, Rebecca Soskin +9 · 90 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  10. Data Decisions and Theoretical Implications when Adversarially Learning Fair Representations
    2017/07/01 by Alex Beutel, Jilin Chen, Beutel, Alex +4 · 17 citations
    Computer Science · Social Sciences · #Adversarial Robustness in Machine Learning #Computers and Society (cs.CY) #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  11. Fairness without Demographics through Adversarially Reweighted Learning
    2020/06/23 by Preethi Lahoti, Alex Beutel, Lahoti, Preethi +13 · 13 citations
    Computer Science · Social Sciences · #Adversarial Robustness in Machine Learning #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  12. Counterfactual Fairness in Text Classification through Robustness
    2018/09/27 by Garg, Sahaj, Perot, Vincent, Limtiaco, Nicole +3 · 9 citations
    #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  13. Controlled Decoding from Language Models
    2023/10/25 by Mudgal, Sidharth, Lee, Jong, Ganapathy, Harish +10 · 17 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  14. Rule Based Rewards for Language Model Safety
    2024/11/02 by Tong Mu, Mu, Tong, Alec Helyar +17 · 28 citations
    Computer Science · Engineering · Social Sciences · #Access Control and Trust #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Safety Systems Engineering in Autonomy #Software Reliability and Analysis Research
  15. Fairness in Recommendation Ranking through Pairwise Comparisons
    2019/03/02 by Beutel, Alex, Chen, Jilin, Doshi, Tulsee +8 · 8 citations
    #Artificial Intelligence (cs.AI) #Computers and Society (cs.CY) #FOS: Computer and information sciences #Information Retrieval (cs.IR) #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  16. Improving Diversity of Demographic Representation in Large Language Models via Collective-Critiques and Self-Voting
    2023/10/25 by Preethi Lahoti, Nicholas Blumm, Lahoti, Preethi +19 · 8 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Speech and dialogue systems #Topic Modeling
  17. Putting Fairness Principles into Practice: Challenges, Metrics, and Improvements
    2019/01/14 by Beutel, Alex, Chen, Jilin, Doshi, Tulsee +6 · 4 citations
    #Artificial Intelligence (cs.AI) #Computers and Society (cs.CY) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  18. Toward a better trade-off between performance and fairness with kernel-based distribution matching
    2019/10/25 by Prost, Flavien, Qian, Hai, Chen, Qiuwen +3 · 3 citations
    #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  19. Measuring Recommender System Effects with Simulated Users
    2021/01/12 by Sirui Yao, Yoni Halpern, Yao, Sirui +15 · 3 citations
    Computer Science · Decision Sciences · Business, Management and Accounting · #Recommender Systems and Techniques #Advanced Bandit Algorithms Research #Consumer Market Behavior and Pricing
  20. Evaluating Fairness of Machine Learning Models Under Uncertain and Incomplete Information
    2021/02/16 by Pranjal Awasthi, Alex Beutel, Awasthi, Pranjal +7 · 3 citations
    Computer Science · Social Sciences · #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Privacy-Preserving Technologies in Data
  21. EdgeCentric: Anomaly Detection in Edge-Attributed Networks
    2015/10/19 by Shah, Neil, Beutel, Alex, Hooi, Bryan +5 · 2 citations
    #FOS: Computer and information sciences #Information Retrieval (cs.IR) #Social and Information Networks (cs.SI)
  22. From Hard Refusals to Safe-Completions: Toward Output-Centric Safety Training
    2025/08/12 by Yuan, Yuan, Sriskandarajah, Tina, Brakman, Anna-Luisa +4 · 17 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computers and Society (cs.CY) #FOS: Computer and information sciences
  23. Diverse and Effective Red Teaming with Auto-generated Rewards and Multi-step Reinforcement Learning
    2024/12/24 by Beutel, Alex, Xiao, Kai, Heidecke, Johannes +1 · 6 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  24. First-Person Fairness in Chatbots
    2024/10/16 by Tyna Eloundou, Alex Beutel, Eloundou, Tyna +17 · 5 citations
    Social Sciences · #Ethics and Social Impacts of AI #Digital Economy and Work Transformation
  25. Spotting Suspicious Link Behavior with fBox: An Adversarial Perspective
    2014/10/15 by Neil Shah, Shah, Neil, Alex Beutel +5 · 1 citation
    Computer Science · #Advanced Malware Detection Techniques #FOS: Computer and information sciences #Information Retrieval (cs.IR) #Machine Learning (cs.LG) #Network Security and Intrusion Detection #Social and Information Networks (cs.SI) #Spam and Phishing Detection
  26. BIRDNEST: Bayesian Inference for Ratings-Fraud Detection
    2015/11/19 by Hooi, Bryan, Shah, Neil, Beutel, Alex +5 · 1 citation
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Social and Information Networks (cs.SI)
  27. Transfer of Machine Learning Fairness across Domains
    2019/06/24 by Schumann, Candice, Wang, Xuezhi, Beutel, Alex +3 · 1 citation
    #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  28. Improving Calibration through the Relationship with Adversarial Robustness
    2020/06/29 by Qin, Yao, Wang, Xuezhi, Beutel, Alex +1 · 1 citation
    #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  29. CAT-Gen: Improving Robustness in NLP Models via Controlled Adversarial Text Generation
    2020/10/05 by Tianlu Wang, Wang, Tianlu, Xuezhi Wang +13 · 1 citation
    Computer Science · #Adversarial Robustness in Machine Learning #Computation and Language (cs.CL) #FOS: Computer and information sciences #Hate Speech and Cyberbullying Detection #Topic Modeling
  30. Understanding and Improving Robustness of Vision Transformers through Patch-based Negative Augmentation
    2021/10/15 by Qin, Yao, Zhang, Chiyuan, Chen, Ting +3 · 1 citation
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  31. What Are Effective Labels for Augmented Data? Improving Calibration and Robustness with AutoLabel
    2023/02/22 by Yao Qin, Xuezhi Wang, Qin, Yao +7 · 1 citation
    Computer Science · #Advanced Neural Network Applications #FOS: Computer and information sciences #Image Enhancement Techniques #Machine Learning (cs.LG) #Machine Learning and Data Classification
  32. Striving for data-model efficiency: Identifying data externalities on group performance
    2022/11/11 by Esther Rolf, Ben Packer, Rolf, Esther +5 · 1 citation
    Computer Science · Decision Sciences · #Adversarial Robustness in Machine Learning #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Forecasting Techniques and Applications #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  33. Effective Robustness against Natural Distribution Shifts for Models with Different Training Data
    2023/02/02 by Zhouxing Shi, Shi, Zhouxing, Nicholas Carlini +11 · 1 citation
    Computer Science · #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimodal Machine Learning Applications
  34. Let's Do a Thought Experiment: Using Counterfactuals to Improve Moral Reasoning
    2023/06/25 by Xiao Ma, Swaroop Mishra, Ma, Xiao +7 · 1 citation
    Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Topic Modeling
  35. Improving Few-shot Generalization of Safety Classifiers via Data Augmented Parameter-Efficient Fine-Tuning
    2023/10/25 by Ananth Balashankar, Xiao Ma, Balashankar, Ananth +11 · 1 citation
    Computer Science · Decision Sciences · Health Professions · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Occupational Health and Safety Research #Risk and Safety Analysis #Software Reliability and Analysis Research