Shen, Xinyue
- "Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models
2023/08/07 by Xinyue Shen, Zeyuan Chen, Shen, Xinyue +7 · 81 citations
Computer Science · #Adversarial Robustness in Machine Learning #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Hate Speech and Cyberbullying Detection #Machine Learning (cs.LG) #Topic Modeling
- Unsafe Diffusion: On the Generation of Unsafe Images and Hateful Memes From Text-To-Image Models
2023/05/23 by Yiting Qu, Xinyue Shen, Qu, Yiting +9 · 22 citations
Computer Science · #Generative Adversarial Networks and Image Synthesis #Digital Media Forensic Detection #Adversarial Robustness in Machine Learning
- Comprehensive Assessment of Jailbreak Attacks Against LLMs
2024/02/08 by Junjie Chu, Chu, Junjie, Yugeng Liu +9 · 25 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #Cybercrime and Law Enforcement Studies #FOS: Computer and information sciences #Information and Cyber Security #Law, AI, and Intellectual Property #Machine Learning (cs.LG)
- Disciplined Convex-Concave Programming
2016/04/10 by Xinyue Shen, Steven Diamond, Shen, Xinyue +5 · 7 citations
Engineering · Mathematics · Computer Science · #Sparse and Compressive Sensing Techniques #Advanced Optimization Algorithms Research #Machine Learning and Algorithms
- Prompt Stealing Attacks Against Text-to-Image Generation Models
2023/02/20 by Shen, Xinyue, Qu, Yiting, Backes, Michael +1 · 11 citations
#Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- UnsafeBench: Benchmarking Image Safety Classifiers on Real-World and AI-Generated Images
2024/05/06 by Qu, Yiting, Shen, Xinyue, Wu, Yixin +3 · 15 citations
#Computer Vision and Pattern Recognition (cs.CV) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Social and Information Networks (cs.SI)
- MGTBench: Benchmarking Machine-Generated Text Detection
2023/03/26 by Xinlei He, He, Xinlei, Xinyue Shen +7 · 7 citations
Computer Science · #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Topic Modeling
- Disciplined Multi-Convex Programming
2016/09/12 by Xinyue Shen, Shen, Xinyue, Steven Diamond +7 · 3 citations
Mathematics · Decision Sciences · Computer Science · #Advanced Optimization Algorithms Research #Multi-Criteria Decision Making #Advanced Multi-Objective Optimization Algorithms
- Are We in the AI-Generated Text World Already? Quantifying and Monitoring AIGT on Social Media
2024/12/24 by Sun, Zhen, Zhang, Zongmin, Shen, Xinyue +5 · 8 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Social and Information Networks (cs.SI)
- Voice Jailbreak Attacks Against GPT-4o
2024/05/29 by Xinyue Shen, Yixin Wu, Shen, Xinyue +5 · 3 citations
Social Sciences · Medicine · Computer Science · #Artificial Intelligence in Law #Artificial Intelligence in Healthcare and Education #Cryptography and Data Security
ModSCAN: Measuring Stereotypical Bias in Large Vision-Language Models from Vision and Language Modalities
2024/10/09 by Jiang, Yukun, Li, Zheng, Shen, Xinyue +3 · 5 citations
#Computers and Society (cs.CY) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences
- HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns
2025/01/28 by Xinyue Shen, Shen, Xinyue, Yixin Wu +9 · 5 citations
Computer Science · #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Hate Speech and Cyberbullying Detection #Machine Learning (cs.LG)
- In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT
2023/04/18 by Xinyue Shen, Shen, Xinyue, Zeyuan Chen +5 · 2 citations
Computer Science · Medicine · #Topic Modeling #Artificial Intelligence in Healthcare and Education #Explainable Artificial Intelligence (XAI)