Deyue Zhang
- Multi-Turn Context Jailbreak Attack on Large Language Models From First Principles
2024/08/08 by Xiongtao Sun, Sun, Xiongtao, Deyue Zhang +7 · 16 citations
Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
- Reasoning-Augmented Conversation for Multi-Turn Jailbreak Attacks on Large Language Models
2025/02/16 by Zonghao Ying, Ying, Zonghao, Deyue Zhang +17 · 22 citations
Computer Science · Psychology · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #Deception detection and forensic psychology #FOS: Computer and information sciences #Hate Speech and Cyberbullying Detection
- Towards Understanding the Safety Boundaries of DeepSeek Models: Evaluation and Findings
2025/03/19 by Zonghao Ying, Ying, Zonghao, Yongxin Huang +14 · 9 citations
Decision Sciences · #Scientific Computing and Data Management
- Dynamic Defense Profiling Enables Cognitive Jailbreak of Text-to-Image Models
2026/07/20 by Dongdong Yang, Deyue Zhang, Zhao Liu +5
#cs.AI