vix.ing · top · new · best · stats · spec

Debenedetti, Edoardo

  1. Defeating Prompt Injections by Design
    2025/03/24 by Edoardo Debenedetti, Ilia Shumailov, Debenedetti, Edoardo +17 · 23 voices · 58 citations
    #cs.CR #cs.AI
  2. Design Patterns for Securing LLM Agents against Prompt Injections
    2025/06/10 by Luca Beurer-Kellner, Beurer-Kellner, Luca, Beat Buesser +25 · 10 voices · 19 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Explainable Artificial Intelligence (XAI) #Topic Modeling
  3. JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models
    2024/03/28 by Patrick Chao, Edoardo Debenedetti, Chao, Patrick +21 · 90 citations
    Computer Science · #Authorship Attribution and Profiling #Cryptography and Security (cs.CR) #Digital and Cyber Forensics #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques
  4. RobustBench: a standardized adversarial robustness benchmark
    2020/10/19 by Croce, Francesco, Andriushchenko, Maksym, Sehwag, Vikash +5 · 46 citations
    #Computer Vision and Pattern Recognition (cs.CV) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  5. AgentDojo-PROV: A W3C PROV-O Corpus of LLM Agent Executions
    2024/06/19 by Debenedetti, Edoardo, Jie Zhang, Zhang, Jie +8 · 19 citations
    Computer Science · #Advanced Malware Detection Techniques #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multi-Agent Systems and Negotiation #Network Security and Intrusion Detection
  6. Adversarial Search Engine Optimization for Large Language Models
    2024/06/26 by Fredrik Nestaas, Edoardo Debenedetti, Nestaas, Fredrik +3 · 3 voices · 11 citations
    #cs.CR #cs.LG
  7. Dataset and Lessons Learned from the 2024 SaTML LLM Capture-the-Flag Competition
    2024/06/12 by Debenedetti, Edoardo, Rando, Javier, Paleka, Daniel +18 · 5 citations
    #Artificial Intelligence (cs.AI) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences
  8. Privacy Side Channels in Machine Learning Systems
    2023/09/11 by Edoardo Debenedetti, Giorgio Severi, Debenedetti, Edoardo +13 · 4 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Cryptography and Data Security #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Privacy-Preserving Technologies in Data
  9. A Light Recipe to Train Robust Vision Transformers
    2022/09/15 by Debenedetti, Edoardo, Sehwag, Vikash, Mittal, Prateek · 3 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  10. Exploring Memorization and Copyright Violation in Frontier LLMs: A Study of the New York Times v. OpenAI 2023 Lawsuit
    2024/12/09 by Joshua Freeman, Chloe Rippe, Freeman, Joshua +5 · 5 citations
    Social Sciences · Business, Management and Accounting · #Legal Systems and Judicial Processes #Intellectual Property Law #Business Law and Ethics
  11. Measuring Non-Adversarial Reproduction of Training Data in Large Language Models
    2024/11/15 by Michael Aerni, Javier Rando, Aerni, Michael +9 · 2 voices · 2 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Topic Modeling #cs.CL #cs.LG
  12. AutoAdvExBench: Benchmarking autonomous exploitation of adversarial example defenses
    2025/03/03 by Nicholas Carlini, Carlini, Nicholas, Javier Rando +7 · 3 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Advanced Malware Detection Techniques #Security and Verification in Computing
  13. Evading Black-box Classifiers Without Breaking Eggs
    2023/06/05 by Debenedetti, Edoardo, Carlini, Nicholas, Tramèr, Florian · 1 citation
    #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  14. AI Risk Management Should Incorporate Both Safety and Security
    2024/05/29 by Xiangyu Qi, Qi, Xiangyu, Yangsibo Huang +47 · 1 citation
    Medicine · Social Sciences · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Healthcare and Education #Cryptography and Security (cs.CR) #Ethics and Social Impacts of AI #FOS: Computer and information sciences