vix.ing · top · new · best · stats · spec

He, Xinlei

  1. Jailbreak Attacks and Defenses Against Large Language Models: A Survey
    2024/07/05 by Sibo Yi, Yule Liu, Yi, Sibo +13 · 75 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Hate Speech and Cyberbullying Detection #Machine Learning (cs.LG) #Natural Language Processing Techniques
  2. Unsafe Diffusion: On the Generation of Unsafe Images and Hateful Memes From Text-To-Image Models
    2023/05/23 by Yiting Qu, Xinyue Shen, Qu, Yiting +9 · 34 citations
    Computer Science · #Generative Adversarial Networks and Image Synthesis #Digital Media Forensic Detection #Adversarial Robustness in Machine Learning
  3. Stealing Links from Graph Neural Networks
    2020/05/05 by Xinlei He, Jinyuan Jia, He, Xinlei +7 · 10 citations
    Computer Science · #Advanced Graph Neural Networks #Adversarial Robustness in Machine Learning #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Privacy-Preserving Technologies in Data
  4. ML-Doctor: Holistic Risk Assessment of Inference Attacks Against Machine Learning Models
    2021/02/04 by Yugeng Liu, Rui Wen, Liu, Yugeng +15 · 11 citations
    Computer Science · Medicine · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Artificial Intelligence in Healthcare and Education #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Privacy-Preserving Technologies in Data
  5. MGTBench: Benchmarking Machine-Generated Text Detection
    2023/03/26 by Xinlei He, Xinyue Shen, He, Xinlei +7 · 13 citations
    Computer Science · #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Topic Modeling
  6. You Only Prompt Once: On the Capabilities of Prompt Learning on Large Language Models to Tackle Toxic Content
    2023/08/10 by Xinlei He, He, Xinlei, Savvas Zannettou +5 · 11 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning and Data Classification #Social and Information Networks (cs.SI)
  7. Data Poisoning Attacks Against Multimodal Encoders
    2022/09/30 by Yang, Ziqing, He, Xinlei, Li, Zheng +4 · 9 citations
    #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  8. Node-Level Membership Inference Attacks Against Graph Neural Networks
    2021/02/10 by Xinlei He, Rui Wen, He, Xinlei +9 · 7 citations
    Computer Science · #Advanced Graph Neural Networks #Adversarial Robustness in Machine Learning #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Privacy-Preserving Technologies in Data
  9. JailbreakEval: An Integrated Toolkit for Evaluating Jailbreak Attempts Against Large Language Models
    2024/06/13 by Ran, Delong, Liu, Jinyuan, Gong, Yichen +4 · 12 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences
  10. Have You Merged My Model? On The Robustness of Large Language Model IP Protection Methods Against Model Merging
    2024/04/08 by Tianshuo Cong, Delong Ran, Cong, Tianshuo +15 · 9 citations
    Social Sciences · #Access Control and Trust #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences
  11. Model Stealing Attacks Against Inductive Graph Neural Networks
    2021/12/15 by Shen, Yun, He, Xinlei, Han, Yufei +1 · 5 citations
    #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  12. Can't Steal? Cont-Steal! Contrastive Stealing Attacks Against Image Encoders
    2022/01/19 by Zeyang Sha, Xinlei He, Sha, Zeyang +7 · 5 citations
    Computer Science · #Adversarial Robustness in Machine Learning
  13. Fine-Tuning Is All You Need to Mitigate Backdoor Attacks
    2022/12/18 by Sha, Zeyang, He, Xinlei, Berrang, Pascal +2 · 6 citations
    #Computer Vision and Pattern Recognition (cs.CV) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  14. Auditing Membership Leakages of Multi-Exit Networks
    2022/08/23 by Zheng Li, Li, Zheng, Yiyong Liu +9 · 6 citations
    Computer Science · #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Network Security and Intrusion Detection #Security in Wireless Sensor Networks #Software-Defined Networks and 5G
  15. FC-Attack: Jailbreaking Multimodal Large Language Models via Auto-Generated Flowcharts
    2025/02/28 by Ziyi Zhang, Zhang, Ziyi, Zhen Sun +7 · 10 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Hate Speech and Cyberbullying Detection #Machine Learning (cs.LG) #Topic Modeling
  16. Are We in the AI-Generated Text World Already? Quantifying and Monitoring AIGT on Social Media
    2024/12/24 by Sun, Zhen, Zhang, Zongmin, Shen, Xinyue +5 · 9 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Social and Information Networks (cs.SI)
  17. SSLGuard: A Watermarking Scheme for Self-supervised Learning Pre-trained Encoders
    2022/01/27 by Cong, Tianshuo, He, Xinlei, Zhang, Yang · 4 citations
    #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  18. Semi-Leak: Membership Inference Attacks Against Semi-supervised Learning
    2022/07/25 by He, Xinlei, Liu, Hongbin, Gong, Neil Zhenqiang +1 · 3 citations
    #Computer Vision and Pattern Recognition (cs.CV) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  19. Unsafe LLM-Based Search: Quantitative Analysis and Mitigation of Safety Risks in AI Web Search
    2025/02/07 by Luo, Zeren, Peng, Zifan, Yule Liu +9 · 6 citations
    Business, Management and Accounting · #Big Data and Business Intelligence
  20. Generative Watermarking Against Unauthorized Subject-Driven Image Synthesis
    2023/06/13 by Yihan Ma, Ma, Yihan, Zhengyu Zhao +9 · 3 citations
    Computer Science · #Advanced Steganography and Watermarking Techniques #Computer Vision and Pattern Recognition (cs.CV) #Cryptography and Security (cs.CR) #Digital Media Forensic Detection #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis
  21. Test-Time Poisoning Attacks Against Test-Time Adaptation Models
    2023/08/16 by Tianshuo Cong, Cong, Tianshuo, Xinlei He +5 · 3 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Cryptography and Security (cs.CR) #FOS: Computer and information sciences
  22. PEFTGuard: Detecting Backdoor Attacks Against Parameter-Efficient Fine-Tuning
    2024/11/26 by Zhen Sun, Tianshuo Cong, Sun, Zhen +13 · 6 citations
    Computer Science · #Advanced Malware Detection Techniques #Cryptographic Implementations and Security #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Security and Verification in Computing
  23. Automatic Dataset Construction (ADC): Sample Collection, Data Curation, and Beyond
    2024/08/21 by Minghao Liu, Liu, Minghao, Zengfeng Di +29 · 7 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Data Analysis with R #FOS: Computer and information sciences #Machine Learning (cs.LG)
  24. FacLens: Transferable Probe for Foreseeing Non-Factuality in Fact-Seeking Question Answering of Large Language Models
    2024/06/08 by Wang, Yanling, Li, Haoyang, Zou, Hao +4 · 3 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  25. SecurityNet: Assessing Machine Learning Vulnerabilities on Public Models
    2023/10/19 by Zhang, Boyang, Li, Zheng, Yang, Ziqing +4 · 3 citations
    #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  26. CL-Attack: Textual Backdoor Attacks via Cross-Lingual Triggers
    2024/12/26 by Jingyi Zheng, Tao Hu, Zheng, Jingyi +5 · 4 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Hate Speech and Cyberbullying Detection
  27. Humanizing LLMs: A Survey of Psychological Measurements with Tools, Datasets, and Human-Agent Applications
    2025/04/30 by Dong, Wenhan, Zhao, Yuemeng, Sun, Zhen +10 · 5 citations
    #Computation and Language (cs.CL) #Computers and Society (cs.CY) #FOS: Computer and information sciences #Human-Computer Interaction (cs.HC) #Machine Learning (cs.LG)
  28. Membership-Doctor: Comprehensive Assessment of Membership Inference Against Machine Learning Models
    2022/08/22 by Xinlei He, He, Xinlei, Zheng Li +6 · 2 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Network Security and Intrusion Detection
  29. On the Evolution of (Hateful) Memes by Means of Multimodal Contrastive Learning
    2022/12/13 by Yiting Qu, Xinlei He, Qu, Yiting +9 · 3 citations
    Computer Science · Psychology · Social Sciences · #Hate Speech and Cyberbullying Detection #Humor Studies and Applications #Misinformation and Its Impacts
  30. Thought Manipulation: External Thought Can Be Efficient for Large Reasoning Models
    2025/04/18 by Liu, Yule, Zheng, Jingyi, Sun, Zhen +6 · 6 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
  31. On Evaluating The Performance of Watermarked Machine-Generated Texts Under Adversarial Attacks
    2024/07/05 by Zesen Liu, Tianshuo Cong, Liu, Zesen +5 · 2 citations
    Computer Science · #Advanced Malware Detection Techniques #Adversarial Robustness in Machine Learning #Network Security and Intrusion Detection
  32. Backdoor Attack on Vision Language Models with Stealthy Semantic Manipulation
    2025/06/08 by Zhiyuan Zhong, Zhen Sun, Zhong, Zhiyuan +7 · 5 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Multimodal Machine Learning Applications #Advanced Neural Network Applications
  33. GUARD: Generation-time LLM Unlearning via Adaptive Restriction and Detection
    2025/05/19 by Zhijie Deng, C. Liu, Deng, Zhijie +13 · 6 citations
    Computer Science · Medicine · #Topic Modeling #Domain Adaptation and Few-Shot Learning #Artificial Intelligence in Healthcare and Education
  34. Quantifying and Mitigating Privacy Risks of Contrastive Learning
    2021/02/08 by Xinlei He, He, Xinlei, Yang Zhang +1 · 1 citation
    Computer Science · #Adversarial Robustness in Machine Learning #Cryptography and Security (cs.CR) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Privacy-Preserving Technologies in Data
  35. JALMBench: Benchmarking Jailbreak Vulnerabilities in Audio Language Models
    2025/05/23 by Peng, Zifan, Liu, Yule, Sun, Zhen +9 · 3 citations
    #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #FOS: Electrical engineering #Sound (cs.SD) #electronic engineering #information engineering
  36. Can VLMs Detect and Localize Fine-Grained AI-Edited Images?
    2025/05/21 by Zhen Sun, Sun, Zhen, Ziyi Zhang +22 · 3 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Multimodal Machine Learning Applications
  37. Membership Inference Attack Against Masked Image Modeling
    2024/08/13 by Zheng Li, Xinlei He, Li, Zheng +5 · 2 citations
    Medicine · Engineering · #Artificial Intelligence in Healthcare and Education #Medical Imaging and Analysis
  38. SoK: Benchmarking Poisoning Attacks and Defenses in Federated Learning
    2025/02/06 by Zhang, Heyi, Liu, Yule, He, Xinlei +3 · 2 citations
    #Artificial Intelligence (cs.AI) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  39. Quantized Delta Weight Is Safety Keeper
    2024/11/29 by Yule Liu, Liu, Yule, Zhen Sun +4 · 2 citations
    Decision Sciences · #Artificial Intelligence (cs.AI) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Risk and Safety Analysis
  40. LoRA-Leak: Membership Inference Attacks Against LoRA Fine-tuned Language Models
    2025/07/24 by Delong Ran, Ran, Delong, Xinlei He +9 · 2 citations
    Computer Science · Medicine · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Healthcare and Education #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Topic Modeling
  41. On the Generalization and Adaptation Ability of Machine-Generated Text Detectors in Academic Writing
    2024/12/23 by Liu, Yule, Zhong, Zhiyuan, Liao, Yifan +8 · 1 citation
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences