Xiong Chen
- Defensive Prompt Patch: A Robust and Interpretable Defense of LLMs against Jailbreak Attacks
2024/05/30 by Xiong Chen, Xiong, Chen, Xiangyu Qi +5 · 11 citations
Computer Science · #Adversarial Robustness in Machine Learning #Cryptography and Data Security #Cryptography and Security (cs.CR) #FOS: Computer and information sciences