Debenedetti, Edoardo
- Defeating Prompt Injections by Design
2025/03/24 by Edoardo Debenedetti, Ilia Shumailov, Debenedetti, Edoardo +17 · 23 voices · 58 citations
#cs.CR #cs.AI
- Design Patterns for Securing LLM Agents against Prompt Injections
2025/06/10 by Luca Beurer-Kellner, Beurer-Kellner, Luca, Beat Buesser +25 · 10 voices · 19 citations
Computer Science · #Adversarial Robustness in Machine Learning #Explainable Artificial Intelligence (XAI) #Topic Modeling
- JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models
2024/03/28 by Patrick Chao, Edoardo Debenedetti, Chao, Patrick +21 · 90 citations
Computer Science · #Authorship Attribution and Profiling #Cryptography and Security (cs.CR) #Digital and Cyber Forensics #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques
- RobustBench: a standardized adversarial robustness benchmark
2020/10/19 by Croce, Francesco, Andriushchenko, Maksym, Sehwag, Vikash +5 · 46 citations
#Computer Vision and Pattern Recognition (cs.CV) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- AgentDojo-PROV: A W3C PROV-O Corpus of LLM Agent Executions
2024/06/19 by Debenedetti, Edoardo, Jie Zhang, Zhang, Jie +8 · 19 citations
Computer Science · #Advanced Malware Detection Techniques #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multi-Agent Systems and Negotiation #Network Security and Intrusion Detection
- Adversarial Search Engine Optimization for Large Language Models
2024/06/26 by Fredrik Nestaas, Edoardo Debenedetti, Nestaas, Fredrik +3 · 3 voices · 11 citations
#cs.CR #cs.LG
- Dataset and Lessons Learned from the 2024 SaTML LLM Capture-the-Flag Competition
2024/06/12 by Debenedetti, Edoardo, Rando, Javier, Paleka, Daniel +18 · 5 citations
#Artificial Intelligence (cs.AI) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences
- Privacy Side Channels in Machine Learning Systems
2023/09/11 by Edoardo Debenedetti, Giorgio Severi, Debenedetti, Edoardo +13 · 4 citations
Computer Science · #Adversarial Robustness in Machine Learning #Cryptography and Data Security #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Privacy-Preserving Technologies in Data
- A Light Recipe to Train Robust Vision Transformers
2022/09/15 by Debenedetti, Edoardo, Sehwag, Vikash, Mittal, Prateek · 3 citations
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Exploring Memorization and Copyright Violation in Frontier LLMs: A Study of the New York Times v. OpenAI 2023 Lawsuit
2024/12/09 by Joshua Freeman, Chloe Rippe, Freeman, Joshua +5 · 5 citations
Social Sciences · Business, Management and Accounting · #Legal Systems and Judicial Processes #Intellectual Property Law #Business Law and Ethics
- Measuring Non-Adversarial Reproduction of Training Data in Large Language Models
2024/11/15 by Michael Aerni, Javier Rando, Aerni, Michael +9 · 2 voices · 2 citations
Computer Science · #Adversarial Robustness in Machine Learning #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Topic Modeling #cs.CL #cs.LG
- AutoAdvExBench: Benchmarking autonomous exploitation of adversarial example defenses
2025/03/03 by Nicholas Carlini, Carlini, Nicholas, Javier Rando +7 · 3 citations
Computer Science · #Adversarial Robustness in Machine Learning #Advanced Malware Detection Techniques #Security and Verification in Computing
- Evading Black-box Classifiers Without Breaking Eggs
2023/06/05 by Debenedetti, Edoardo, Carlini, Nicholas, Tramèr, Florian · 1 citation
#Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- AI Risk Management Should Incorporate Both Safety and Security
2024/05/29 by Xiangyu Qi, Qi, Xiangyu, Yangsibo Huang +47 · 1 citation
Medicine · Social Sciences · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Healthcare and Education #Cryptography and Security (cs.CR) #Ethics and Social Impacts of AI #FOS: Computer and information sciences