He, Xinlei
- Jailbreak Attacks and Defenses Against Large Language Models: A Survey
2024/07/05 by Sibo Yi, Yule Liu, Yi, Sibo +13 · 75 citations
Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Hate Speech and Cyberbullying Detection #Machine Learning (cs.LG) #Natural Language Processing Techniques
- Unsafe Diffusion: On the Generation of Unsafe Images and Hateful Memes From Text-To-Image Models
2023/05/23 by Yiting Qu, Xinyue Shen, Qu, Yiting +9 · 34 citations
Computer Science · #Generative Adversarial Networks and Image Synthesis #Digital Media Forensic Detection #Adversarial Robustness in Machine Learning
- Stealing Links from Graph Neural Networks
2020/05/05 by Xinlei He, Jinyuan Jia, He, Xinlei +7 · 10 citations
Computer Science · #Advanced Graph Neural Networks #Adversarial Robustness in Machine Learning #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Privacy-Preserving Technologies in Data
- ML-Doctor: Holistic Risk Assessment of Inference Attacks Against Machine Learning Models
2021/02/04 by Yugeng Liu, Rui Wen, Liu, Yugeng +15 · 11 citations
Computer Science · Medicine · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Artificial Intelligence in Healthcare and Education #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Privacy-Preserving Technologies in Data
- MGTBench: Benchmarking Machine-Generated Text Detection
2023/03/26 by Xinlei He, Xinyue Shen, He, Xinlei +7 · 13 citations
Computer Science · #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Topic Modeling
- You Only Prompt Once: On the Capabilities of Prompt Learning on Large Language Models to Tackle Toxic Content
2023/08/10 by Xinlei He, He, Xinlei, Savvas Zannettou +5 · 11 citations
Computer Science · #Adversarial Robustness in Machine Learning #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning and Data Classification #Social and Information Networks (cs.SI)
- Data Poisoning Attacks Against Multimodal Encoders
2022/09/30 by Yang, Ziqing, He, Xinlei, Li, Zheng +4 · 9 citations
#Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Node-Level Membership Inference Attacks Against Graph Neural Networks
2021/02/10 by Xinlei He, Rui Wen, He, Xinlei +9 · 7 citations
Computer Science · #Advanced Graph Neural Networks #Adversarial Robustness in Machine Learning #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Privacy-Preserving Technologies in Data
- JailbreakEval: An Integrated Toolkit for Evaluating Jailbreak Attempts Against Large Language Models
2024/06/13 by Ran, Delong, Liu, Jinyuan, Gong, Yichen +4 · 12 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences
- Have You Merged My Model? On The Robustness of Large Language Model IP Protection Methods Against Model Merging
2024/04/08 by Tianshuo Cong, Delong Ran, Cong, Tianshuo +15 · 9 citations
Social Sciences · #Access Control and Trust #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences
- Model Stealing Attacks Against Inductive Graph Neural Networks
2021/12/15 by Shen, Yun, He, Xinlei, Han, Yufei +1 · 5 citations
#Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Can't Steal? Cont-Steal! Contrastive Stealing Attacks Against Image Encoders
2022/01/19 by Zeyang Sha, Xinlei He, Sha, Zeyang +7 · 5 citations
Computer Science · #Adversarial Robustness in Machine Learning
- Fine-Tuning Is All You Need to Mitigate Backdoor Attacks
2022/12/18 by Sha, Zeyang, He, Xinlei, Berrang, Pascal +2 · 6 citations
#Computer Vision and Pattern Recognition (cs.CV) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Auditing Membership Leakages of Multi-Exit Networks
2022/08/23 by Zheng Li, Li, Zheng, Yiyong Liu +9 · 6 citations
Computer Science · #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Network Security and Intrusion Detection #Security in Wireless Sensor Networks #Software-Defined Networks and 5G
- FC-Attack: Jailbreaking Multimodal Large Language Models via Auto-Generated Flowcharts
2025/02/28 by Ziyi Zhang, Zhang, Ziyi, Zhen Sun +7 · 10 citations
Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Hate Speech and Cyberbullying Detection #Machine Learning (cs.LG) #Topic Modeling
- Are We in the AI-Generated Text World Already? Quantifying and Monitoring AIGT on Social Media
2024/12/24 by Sun, Zhen, Zhang, Zongmin, Shen, Xinyue +5 · 9 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Social and Information Networks (cs.SI)
- SSLGuard: A Watermarking Scheme for Self-supervised Learning Pre-trained Encoders
2022/01/27 by Cong, Tianshuo, He, Xinlei, Zhang, Yang · 4 citations
#Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Semi-Leak: Membership Inference Attacks Against Semi-supervised Learning
2022/07/25 by He, Xinlei, Liu, Hongbin, Gong, Neil Zhenqiang +1 · 3 citations
#Computer Vision and Pattern Recognition (cs.CV) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Unsafe LLM-Based Search: Quantitative Analysis and Mitigation of Safety Risks in AI Web Search
2025/02/07 by Luo, Zeren, Peng, Zifan, Yule Liu +9 · 6 citations
Business, Management and Accounting · #Big Data and Business Intelligence
- Generative Watermarking Against Unauthorized Subject-Driven Image Synthesis
2023/06/13 by Yihan Ma, Ma, Yihan, Zhengyu Zhao +9 · 3 citations
Computer Science · #Advanced Steganography and Watermarking Techniques #Computer Vision and Pattern Recognition (cs.CV) #Cryptography and Security (cs.CR) #Digital Media Forensic Detection #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis
- Test-Time Poisoning Attacks Against Test-Time Adaptation Models
2023/08/16 by Tianshuo Cong, Cong, Tianshuo, Xinlei He +5 · 3 citations
Computer Science · #Adversarial Robustness in Machine Learning #Cryptography and Security (cs.CR) #FOS: Computer and information sciences
- PEFTGuard: Detecting Backdoor Attacks Against Parameter-Efficient Fine-Tuning
2024/11/26 by Zhen Sun, Tianshuo Cong, Sun, Zhen +13 · 6 citations
Computer Science · #Advanced Malware Detection Techniques #Cryptographic Implementations and Security #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Security and Verification in Computing
- Automatic Dataset Construction (ADC): Sample Collection, Data Curation, and Beyond
2024/08/21 by Minghao Liu, Liu, Minghao, Zengfeng Di +29 · 7 citations
Computer Science · #Artificial Intelligence (cs.AI) #Data Analysis with R #FOS: Computer and information sciences #Machine Learning (cs.LG)
- FacLens: Transferable Probe for Foreseeing Non-Factuality in Fact-Seeking Question Answering of Large Language Models
2024/06/08 by Wang, Yanling, Li, Haoyang, Zou, Hao +4 · 3 citations
#Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- SecurityNet: Assessing Machine Learning Vulnerabilities on Public Models
2023/10/19 by Zhang, Boyang, Li, Zheng, Yang, Ziqing +4 · 3 citations
#Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- CL-Attack: Textual Backdoor Attacks via Cross-Lingual Triggers
2024/12/26 by Jingyi Zheng, Tao Hu, Zheng, Jingyi +5 · 4 citations
Computer Science · #Artificial Intelligence (cs.AI) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Hate Speech and Cyberbullying Detection
- Humanizing LLMs: A Survey of Psychological Measurements with Tools, Datasets, and Human-Agent Applications
2025/04/30 by Dong, Wenhan, Zhao, Yuemeng, Sun, Zhen +10 · 5 citations
#Computation and Language (cs.CL) #Computers and Society (cs.CY) #FOS: Computer and information sciences #Human-Computer Interaction (cs.HC) #Machine Learning (cs.LG)
- Membership-Doctor: Comprehensive Assessment of Membership Inference Against Machine Learning Models
2022/08/22 by Xinlei He, He, Xinlei, Zheng Li +6 · 2 citations
Computer Science · #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Network Security and Intrusion Detection
- On the Evolution of (Hateful) Memes by Means of Multimodal Contrastive Learning
2022/12/13 by Yiting Qu, Xinlei He, Qu, Yiting +9 · 3 citations
Computer Science · Psychology · Social Sciences · #Hate Speech and Cyberbullying Detection #Humor Studies and Applications #Misinformation and Its Impacts
- Thought Manipulation: External Thought Can Be Efficient for Large Reasoning Models
2025/04/18 by Liu, Yule, Zheng, Jingyi, Sun, Zhen +6 · 6 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
- On Evaluating The Performance of Watermarked Machine-Generated Texts Under Adversarial Attacks
2024/07/05 by Zesen Liu, Tianshuo Cong, Liu, Zesen +5 · 2 citations
Computer Science · #Advanced Malware Detection Techniques #Adversarial Robustness in Machine Learning #Network Security and Intrusion Detection
- Backdoor Attack on Vision Language Models with Stealthy Semantic Manipulation
2025/06/08 by Zhiyuan Zhong, Zhen Sun, Zhong, Zhiyuan +7 · 5 citations
Computer Science · #Adversarial Robustness in Machine Learning #Multimodal Machine Learning Applications #Advanced Neural Network Applications
- GUARD: Generation-time LLM Unlearning via Adaptive Restriction and Detection
2025/05/19 by Zhijie Deng, C. Liu, Deng, Zhijie +13 · 6 citations
Computer Science · Medicine · #Topic Modeling #Domain Adaptation and Few-Shot Learning #Artificial Intelligence in Healthcare and Education
- Quantifying and Mitigating Privacy Risks of Contrastive Learning
2021/02/08 by Xinlei He, He, Xinlei, Yang Zhang +1 · 1 citation
Computer Science · #Adversarial Robustness in Machine Learning #Cryptography and Security (cs.CR) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Privacy-Preserving Technologies in Data
- JALMBench: Benchmarking Jailbreak Vulnerabilities in Audio Language Models
2025/05/23 by Peng, Zifan, Liu, Yule, Sun, Zhen +9 · 3 citations
#Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
- Can VLMs Detect and Localize Fine-Grained AI-Edited Images?
2025/05/21 by Zhen Sun, Sun, Zhen, Ziyi Zhang +22 · 3 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Multimodal Machine Learning Applications
- Membership Inference Attack Against Masked Image Modeling
2024/08/13 by Zheng Li, Xinlei He, Li, Zheng +5 · 2 citations
Medicine · Engineering · #Artificial Intelligence in Healthcare and Education #Medical Imaging and Analysis
- SoK: Benchmarking Poisoning Attacks and Defenses in Federated Learning
2025/02/06 by Zhang, Heyi, Liu, Yule, He, Xinlei +3 · 2 citations
#Artificial Intelligence (cs.AI) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Quantized Delta Weight Is Safety Keeper
2024/11/29 by Yule Liu, Liu, Yule, Zhen Sun +4 · 2 citations
Decision Sciences · #Artificial Intelligence (cs.AI) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Risk and Safety Analysis
- LoRA-Leak: Membership Inference Attacks Against LoRA Fine-tuned Language Models
2025/07/24 by Delong Ran, Ran, Delong, Xinlei He +9 · 2 citations
Computer Science · Medicine · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Healthcare and Education #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Topic Modeling
- On the Generalization and Adaptation Ability of Machine-Generated Text Detectors in Academic Writing
2024/12/23 by Liu, Yule, Zhong, Zhiyuan, Liao, Yifan +8 · 1 citation
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences