Lutz, Roman
- Lessons From Red Teaming 100 Generative AI Products
2025/01/13 by Blake Bullwinkel, Bullwinkel, Blake, Amanda Minnich +49 · 16 voices · 9 citations
#cs.AI
- Fairlearn: Assessing and Improving Fairness of AI Systems
2023/03/29 by Hilde Weerts, Miroslav Dudı́k, Weerts, Hilde +9 · 19 citations
Social Sciences · #Artificial Intelligence (cs.AI) #Computers and Society (cs.CY) #Ethics and Social Impacts of AI #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Phi-3 Safety Post-Training: Aligning Language Models with a "Break-Fix" Cycle
2024/07/18 by Haider, Emman, Perez-Becker, Daniel, Portet, Thomas +28 · 6 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
- PyRIT: A Framework for Security Risk Identification and Red Teaming in Generative AI System
2024/10/01 by Munoz, Gary D. Lopez, Minnich, Amanda J., Lutz, Roman +17 · 3 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences