vix.ing · top · new · best · stats · spec

Lutz, Roman

  1. Lessons From Red Teaming 100 Generative AI Products
    2025/01/13 by Blake Bullwinkel, Bullwinkel, Blake, Amanda Minnich +49 · 16 voices · 9 citations
    #cs.AI
  2. Fairlearn: Assessing and Improving Fairness of AI Systems
    2023/03/29 by Hilde Weerts, Miroslav Dudı́k, Weerts, Hilde +9 · 19 citations
    Social Sciences · #Artificial Intelligence (cs.AI) #Computers and Society (cs.CY) #Ethics and Social Impacts of AI #FOS: Computer and information sciences #Machine Learning (cs.LG)
  3. Phi-3 Safety Post-Training: Aligning Language Models with a "Break-Fix" Cycle
    2024/07/18 by Haider, Emman, Perez-Becker, Daniel, Portet, Thomas +28 · 6 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
  4. PyRIT: A Framework for Security Risk Identification and Red Teaming in Generative AI System
    2024/10/01 by Munoz, Gary D. Lopez, Minnich, Amanda J., Lutz, Roman +17 · 3 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences