Chaowei Xiao
- Robust Physical-World Attacks on Deep Learning Models
2017/07/27 by Kevin Eykholt, Eykholt, Kevin, Ivan Evtimov +15 · 3 voices · 1 citation
#cs.CR #cs.LG
- AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models
2023/10/03 by Xiaogeng Liu, Nan Xu, Liu, Xiaogeng +5 · 143 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Text Readability and Simplification #Topic Modeling
- Adversarial Objects Against LiDAR-Based Autonomous Driving Systems
2019/07/11 by Yulong Cao, Chaowei Xiao, Cao, Yulong +11 · 2 voices · 11 citations
Biochemistry, Genetics and Molecular Biology · Computer Science · Mathematics · #Adversarial Robustness in Machine Learning #Bacillus and Francisella bacterial research #Computer Vision and Pattern Recognition (cs.CV) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Forensic and Genetic Research #Machine Learning (cs.LG) #Machine Learning (stat.ML) #cs.CR #cs.CV #cs.LG #stat.ML
- Diffusion Models for Adversarial Purification
2022/05/16 by Weili Nie, Brandon Guo, Nie, Weili +9 · 40 citations
Biochemistry, Genetics and Molecular Biology · Computer Science · #Adversarial Robustness in Machine Learning #Bacillus and Francisella bacterial research #Computer Vision and Pattern Recognition (cs.CV) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG)
- VoxFormer: Sparse Voxel Transformer for Camera-based 3D Semantic Scene Completion
2023/02/23 by Yiming Li, Li, Yiming, Zhiding Yu +13 · 40 citations
Computer Science · Engineering · #3D Shape Modeling and Analysis #Advanced Vision and Imaging #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Robotics (cs.RO) #Robotics and Sensor-Based Localization
- JailBreakV: A Benchmark for Assessing the Robustness of MultiModal Large Language Models against Jailbreak Attacks
2024/04/03 by Weidi Luo, Luo, Weidi, Siyuan Ma +7 · 46 citations
Computer Science · #Hate Speech and Cyberbullying Detection #Digital and Cyber Forensics #Cybercrime and Law Enforcement Studies
- Dolphins: Multimodal Language Model for Driving
2023/12/01 by Yingzi Ma, Yulong Cao, Ma, Yingzi +7 · 32 citations
Computer Science · #Multimodal Machine Learning Applications #Advanced Neural Network Applications #Domain Adaptation and Few-Shot Learning
- Automatic and Universal Prompt Injection Attacks against Large Language Models
2024/03/07 by Xiaogeng Liu, Liu, Xiaogeng, Zhiyuan Yu +7 · 35 citations
Computer Science · #Adversarial Robustness in Machine Learning #Topic Modeling
- AutoDAN-Turbo: A Lifelong Agent for Strategy Self-Exploration to Jailbreak LLMs
2024/10/03 by Xiaogeng Liu, Liu, Xiaogeng, Peiran Li +17 · 44 citations
Computer Science · #Digital and Cyber Forensics #Cybercrime and Law Enforcement Studies #Advanced Malware Detection Techniques
- TrustLLM: Trustworthiness in Large Language Models
2024/01/01 by Yue Huang, Lichao Sun, Huang, Yue +136 · 30 citations
Medicine · Computer Science · #Artificial Intelligence in Healthcare and Education #Privacy-Preserving Technologies in Data #Explainable Artificial Intelligence (XAI)
- Generating Adversarial Examples with Adversarial Networks
2018/01/08 by Chaowei Xiao, Xiao, Chaowei, Bo Li +9 · 18 citations
Computer Science · Physics and Astronomy · #Advanced Malware Detection Techniques #Adversarial Robustness in Machine Learning #Computer Vision and Pattern Recognition (cs.CV) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (stat.ML) #Model Reduction and Neural Networks
- AdaShield: Safeguarding Multimodal Large Language Models from Structure-based Attack via Adaptive Shield Prompting
2024/03/14 by Yu Wang, Wang, Yu, Xiaogeng Liu +7 · 29 citations
Computer Science · #Adversarial Robustness in Machine Learning #Topic Modeling
- EIA: Environmental Injection Attack on Generalist Web Agents for Privacy Leakage
2024/09/17 by Zeyi Liao, Lingbo Mo, Liao, Zeyi +15 · 34 citations
Computer Science · Social Sciences · #Network Security and Intrusion Detection #Access Control and Trust #Spam and Phishing Detection
- Spatially Transformed Adversarial Examples
2018/01/08 by Chaowei Xiao, Jun-Yan Zhu, Xiao, Chaowei +9 · 15 citations
Computer Science · #Advanced Malware Detection Techniques #Computer Vision and Pattern Recognition (cs.CV) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (stat.ML) #Physical Unclonable Functions (PUFs) and Hardware Security
- HaloScope: Harnessing Unlabeled LLM Generations for Hallucination Detection
2024/09/26 by Xuefeng Du, Chaowei Xiao, Du, Xuefeng +3 · 24 citations
Biochemistry, Genetics and Molecular Biology · Engineering · #Cell Image Analysis Techniques #Computation and Language (cs.CL) #FOS: Computer and information sciences #Innovative Microfluidic and Catalytic Techniques Innovation #Machine Learning (cs.LG)
- System-Level Defense against Indirect Prompt Injection Attacks: An Information Flow Control Perspective
2024/09/27 by Fangzhou Wu, Ethan Cecchetti, Wu, Fangzhou +3 · 21 citations
Computer Science · Engineering · #Security and Verification in Computing #Smart Grid Security and Resilience #Advanced Malware Detection Techniques
- Instructions as Backdoors: Backdoor Vulnerabilities of Instruction Tuning for Large Language Models
2023/05/24 by Jiashu Xu, Xu, Jiashu, Mingyu Derek Ma +7 · 11 citations
Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Security and Verification in Computing #Software Testing and Debugging Techniques
- Instructional Fingerprinting of Large Language Models
2024/01/21 by Jiashu Xu, Xu, Jiashu, Fei Wang +9 · 13 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Software Engineering Research #Topic Modeling
- Cognitive Overload: Jailbreaking Large Language Models with Overloaded Logical Thinking
2023/11/16 by Nan Xu, Fei Wang, Xu, Nan +9 · 12 citations
Computer Science · Medicine · #Topic Modeling #Artificial Intelligence in Healthcare and Education #Adversarial Robustness in Machine Learning
- Benchmarking Robustness of 3D Point Cloud Recognition Against Common Corruptions
2022/01/28 by J. F. Sun, Sun, Jiachen, Qingzhao Zhang +9 · 8 citations
Computer Science · Earth and Planetary Sciences · Engineering · #3D Surveying and Cultural Heritage #Advanced Neural Network Applications #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Industrial Vision Systems and Defect Detection #Machine Learning (cs.LG)
- Detecting Backdoors During the Inference Stage Based on Corruption Robustness Consistency
2023/03/27 by Xiaogeng Liu, Liu, Xiaogeng, Minghui Li +13 · 9 citations
Computer Science · #Advanced Neural Network Applications #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #Artificial Intelligence (cs.AI) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- A New Era in LLM Security: Exploring Security Concerns in Real-World LLM-based Systems
2024/02/28 by Fangzhou Wu, Ning Zhang, Wu, Fangzhou +7 · 12 citations
Computer Science · #Blockchain Technology Applications and Security #Digital Rights Management and Security
- Understanding The Robustness in Vision Transformers
2022/04/26 by Daquan Zhou, Zhou, Daquan, Zhiding Yu +11 · 7 citations
Computer Science · #Advanced Neural Network Applications #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Visual Attention and Saliency Detection
- Long-Short Transformer: Efficient Transformers for Language and Vision
2021/07/05 by Chen Zhu, Wei Ping, Zhu, Chen +11 · 7 citations
Computer Science · #Multimodal Machine Learning Applications #Topic Modeling #Domain Adaptation and Few-Shot Learning
- Re-ViLM: Retrieval-Augmented Visual Language Model for Zero and Few-Shot Image Captioning
2023/02/09 by Zhuolin Yang, Yang, Zhuolin, Ping, Wei +28 · 6 citations
Computer Science · #Advanced Image and Video Retrieval Techniques #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Information Retrieval (cs.IR) #Machine Learning (cs.LG) #Multimodal Machine Learning Applications
- Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety
2025/02/02 by Xingjun Ma, Yifeng Gao, Ma, Xingjun +88 · 20 citations
Health Professions · Decision Sciences · Engineering · #Occupational Health and Safety Research #Risk and Safety Analysis #Safety Systems Engineering in Autonomy
- Shall We Pretrain Autoregressive Language Models with Retrieval? A Comprehensive Study
2023/04/13 by Boxin Wang, Wang, Boxin, Peng Xu +20 · 5 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Information Retrieval (cs.IR) #Machine Learning (cs.LG) #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling
- WIPI: A New Web Threat for LLM-Driven Web Agents
2024/02/26 by Fangzhou Wu, Wu, Fangzhou, Shutong Wu +5 · 6 citations
Computer Science · #Artificial Intelligence (cs.AI) #Cryptography and Security (cs.CR) #Digital Rights Management and Security #FOS: Computer and information sciences #Network Security and Intrusion Detection #Web Application Security Vulnerabilities
- Robust Trajectory Prediction against Adversarial Attacks
2022/07/29 by Yulong Cao, Cao, Yulong, Danfei Xu +11 · 4 citations
Computer Science · Engineering · Medicine · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Autonomous Vehicle Technology and Safety #Autopsy Techniques and Outcomes #Computer Vision and Pattern Recognition (cs.CV) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- RLHFPoison: Reward Poisoning Attack for Reinforcement Learning with Human Feedback in Large Language Models
2023/11/16 by Jiongxiao Wang, Wang, Jiongxiao, Junlin Wu +7 · 5 citations
Computer Science · #Topic Modeling #Adversarial Robustness in Machine Learning #Hate Speech and Cyberbullying Detection
- Safeguarding Vision-Language Models Against Patched Visual Prompt Injectors
2024/05/17 by Jiachen Sun, Sun, Jiachen, Changsheng Wang +7 · 6 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #I.2.7 #I.4 #Security and Verification in Computing
- PerAda: Parameter-Efficient Federated Learning Personalization with Generalization Guarantees
2023/02/13 by Chulin Xie, De-An Huang, Xie, Chulin +11 · 4 citations
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Privacy-Preserving Technologies in Data #Recommender Systems and Techniques #Stochastic Gradient Optimization Techniques
- DeceptPrompt: Exploiting LLM-driven Code Generation via Adversarial Natural Language Instructions
2023/12/07 by Fangzhou Wu, Xiaogeng Liu, Wu, Fangzhou +3 · 5 citations
Computer Science · Engineering · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Ferroelectric and Negative Capacitance Devices #Software Engineering Research
- Taxonomy of Machine Learning Safety: A Survey and Primer
2021/06/09 by Sina Mohseni, Mohseni, Sina, Haotao Wang +9 · 3 citations
Computer Science · Engineering · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Safety Systems Engineering in Autonomy #Software Reliability and Analysis Research
- FATH: Authentication-based Test-time Defense against Indirect Prompt Injection Attacks
2024/10/28 by Jiongxiao Wang, Wang, Jiongxiao, Fangzhou Wu +12 · 7 citations
Computer Science · #Advanced Malware Detection Techniques #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Security and Verification in Computing #Software Testing and Debugging Techniques
- Characterizing Attacks on Deep Reinforcement Learning
2019/07/21 by Xinlei Pan, Chaowei Xiao, Pan, Xinlei +17 · 4 citations
Computer Science · #Advanced Malware Detection Techniques #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
- ChatGPT as an Attack Tool: Stealthy Textual Backdoor Attack via Blackbox Generative Model Trigger
2023/04/27 by Jiazhao Li, Yijin Yang, Li, Jiazhao +7 · 4 citations
Computer Science · #Advanced Malware Detection Techniques #Adversarial Robustness in Machine Learning #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Hate Speech and Cyberbullying Detection #Machine Learning (cs.LG)
- LeanAgent: Lifelong Learning for Formal Theorem Proving
2024/10/08 by Adarsh Kumarappan, Mo Tiwari, Kumarappan, Adarsh +10 · 3 voices · 3 citations
Business, Management and Accounting · Computer Science · Engineering · #Artificial Intelligence (cs.AI) #Business Process Modeling and Analysis #FOS: Computer and information sciences #Logic in Computer Science (cs.LO) #Machine Learning (cs.LG) #Manufacturing Process and Optimization #cs.AI #cs.LG #cs.LO
- T-Stitch: Accelerating Sampling in Pre-Trained Diffusion Models with Trajectory Stitching
2024/02/21 by Zizheng Pan, Pan, Zizheng, Bohan Zhuang +13 · 4 citations
Mathematics · Medicine · #Advanced Neuroimaging Techniques and Applications #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Markov Chains and Monte Carlo Methods
- Prismer: A Vision-Language Model with Multi-Task Experts
2023/03/04 by Shikun Liu, Linxi Fan, Liu, Shikun +9 · 3 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimodal Machine Learning Applications #Topic Modeling
- DataGen: Unified Synthetic Dataset Generation via Large Language Models
2024/06/27 by Yi Huang, Huang, Yue, Siyuan Wu +19 · 5 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Topic Modeling
- Exploring the Limits of Domain-Adaptive Training for Detoxifying Large-Scale Language Models
2022/02/08 by Boxin Wang, Wang, Boxin, Ping, Wei +14 · 2 citations
Computer Science · #Topic Modeling #Natural Language Processing Techniques #Adversarial Robustness in Machine Learning
- Mitigating Backdoor Threats to Large Language Models: Advancement and Challenges
2024/09/30 by Qin Liu, Liu, Qin, Wenjie Mo +11 · 4 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Systems and Control (eess.SY) #Topic Modeling #electronic engineering #information engineering
- RePD: Defending Jailbreak Attack through a Retrieval-based Prompt Decomposition Process
2024/10/11 by Peiran Wang, Wang, Peiran, Xiaogeng Liu +3 · 4 citations
Computer Science · #Advanced Malware Detection Techniques #Artificial Intelligence (cs.AI) #Cryptography and Security (cs.CR) #Digital and Cyber Forensics #FOS: Computer and information sciences #Information and Cyber Security
- DRIFT: Dynamic Rule-Based Defense with Injection Isolation for Securing LLM Agents
2025/06/13 by Hao Li, Li, Hao, Xiaogeng Liu +8 · 7 citations
Engineering · Computer Science · #Smart Grid Security and Resilience #Security and Verification in Computing #Network Security and Intrusion Detection
- VLM-Guard: Safeguarding Vision-Language Models via Fulfilling Safety Alignment Gap
2025/02/14 by Liu Qin, Wang Fei, Liu, Qin +5 · 4 citations
Computer Science · #Adversarial Robustness in Machine Learning
- DiffSmooth: Certifiably Robust Learning via Diffusion Models and Local Smoothing
2023/08/28 by Jiawei Zhang, Zhang, Jiawei, Zhong‐Zhu Chen +7 · 2 citations
Computer Science · Medicine · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #COVID-19 diagnosis using AI #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Reinforcement Learning with Human Feedback for Realistic Traffic Simulation
2023/09/01 by Yulong Cao, Boris Ivanovic, Cao, Yulong +5 · 2 citations
Engineering · Psychology · #Traffic control and management #Autonomous Vehicle Technology and Safety #Human-Automation Interaction and Safety
- Semantic Adversarial Attacks via Diffusion Models
2023/09/14 by Chenan Wang, Jinhao Duan, Wang, Chenan +9 · 2 citations
Computer Science · Medicine · #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #Artificial Intelligence (cs.AI) #COVID-19 diagnosis using AI #Computer Vision and Pattern Recognition (cs.CV) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Retrieval-based Controllable Molecule Generation
2022/08/23 by Zichao Wang, Weili Nie, Wang, Zichao +9 · 1 citation
Biochemistry, Genetics and Molecular Biology · Computer Science · Materials Science · #Computational Drug Discovery Methods #FOS: Biological sciences #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning in Materials Science #Protein Structure and Dynamics #Quantitative Methods (q-bio.QM)
- SecretGen: Privacy Recovery on Pre-Trained Models via Distribution Discrimination
2022/07/25 by Zhuowen Yuan, Yuan, Zhuowen, Fan Wu +7 · 1 citation
Computer Science · #Adversarial Robustness in Machine Learning #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Privacy-Preserving Technologies in Data
- Defending against Adversarial Audio via Diffusion Model
2023/03/02 by Shutong Wu, Jiongxiao Wang, Wu, Shutong +6 · 1 citation
Computer Science · #Adversarial Robustness in Machine Learning #Audio and Speech Processing (eess.AS) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #Speech Recognition and Synthesis #electronic engineering #information engineering
- Preference Poisoning Attacks on Reward Model Learning
2024/02/02 by Junlin Wu, Wu, Junlin, Jiongxiao Wang +9 · 1 citation
Pharmacology, Toxicology and Pharmaceutics · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Pharmacovigilance and Adverse Drug Reactions
- A Trembling House of Cards? Mapping Adversarial Attacks against Language Agents
2024/02/15 by Lingbo Mo, Mo, Lingbo, Zeyi Liao +9 · 1 citation
Arts and Humanities · Computer Science · Health Professions · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Interpreting and Communication in Healthcare #Natural Language Processing Techniques #Translation Studies and Practices
- AI Risk Management Should Incorporate Both Safety and Security
2024/05/29 by Xiangyu Qi, Yangsibo Huang, Qi, Xiangyu +47 · 1 citation
Medicine · Social Sciences · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Healthcare and Education #Cryptography and Security (cs.CR) #Ethics and Social Impacts of AI #FOS: Computer and information sciences
- Mitigating Indirect Prompt Injection via Instruction-Following Intent Analysis
2025/11/30 by Mintong Kang, Kang, Mintong, Chong Xiang +9 · 1 citation
Computer Science · #Adversarial Robustness in Machine Learning #Security and Verification in Computing #Topic Modeling