vix.ing · top · new · best · stats · spec

Shen, Xinyue

  1. "Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models
    2023/08/07 by Xinyue Shen, Zeyuan Chen, Shen, Xinyue +7 · 81 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Hate Speech and Cyberbullying Detection #Machine Learning (cs.LG) #Topic Modeling
  2. Unsafe Diffusion: On the Generation of Unsafe Images and Hateful Memes From Text-To-Image Models
    2023/05/23 by Yiting Qu, Xinyue Shen, Qu, Yiting +9 · 22 citations
    Computer Science · #Generative Adversarial Networks and Image Synthesis #Digital Media Forensic Detection #Adversarial Robustness in Machine Learning
  3. Comprehensive Assessment of Jailbreak Attacks Against LLMs
    2024/02/08 by Junjie Chu, Chu, Junjie, Yugeng Liu +9 · 25 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #Cybercrime and Law Enforcement Studies #FOS: Computer and information sciences #Information and Cyber Security #Law, AI, and Intellectual Property #Machine Learning (cs.LG)
  4. Disciplined Convex-Concave Programming
    2016/04/10 by Xinyue Shen, Steven Diamond, Shen, Xinyue +5 · 7 citations
    Engineering · Mathematics · Computer Science · #Sparse and Compressive Sensing Techniques #Advanced Optimization Algorithms Research #Machine Learning and Algorithms
  5. Prompt Stealing Attacks Against Text-to-Image Generation Models
    2023/02/20 by Shen, Xinyue, Qu, Yiting, Backes, Michael +1 · 11 citations
    #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  6. UnsafeBench: Benchmarking Image Safety Classifiers on Real-World and AI-Generated Images
    2024/05/06 by Qu, Yiting, Shen, Xinyue, Wu, Yixin +3 · 15 citations
    #Computer Vision and Pattern Recognition (cs.CV) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Social and Information Networks (cs.SI)
  7. MGTBench: Benchmarking Machine-Generated Text Detection
    2023/03/26 by Xinlei He, He, Xinlei, Xinyue Shen +7 · 7 citations
    Computer Science · #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Topic Modeling
  8. Disciplined Multi-Convex Programming
    2016/09/12 by Xinyue Shen, Shen, Xinyue, Steven Diamond +7 · 3 citations
    Mathematics · Decision Sciences · Computer Science · #Advanced Optimization Algorithms Research #Multi-Criteria Decision Making #Advanced Multi-Objective Optimization Algorithms
  9. Are We in the AI-Generated Text World Already? Quantifying and Monitoring AIGT on Social Media
    2024/12/24 by Sun, Zhen, Zhang, Zongmin, Shen, Xinyue +5 · 8 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Social and Information Networks (cs.SI)
  10. Voice Jailbreak Attacks Against GPT-4o
    2024/05/29 by Xinyue Shen, Yixin Wu, Shen, Xinyue +5 · 3 citations
    Social Sciences · Medicine · Computer Science · #Artificial Intelligence in Law #Artificial Intelligence in Healthcare and Education #Cryptography and Data Security
  11. ModSCAN: Measuring Stereotypical Bias in Large Vision-Language Models from Vision and Language Modalities
    2024/10/09 by Jiang, Yukun, Li, Zheng, Shen, Xinyue +3 · 5 citations
    #Computers and Society (cs.CY) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences
  12. HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns
    2025/01/28 by Xinyue Shen, Shen, Xinyue, Yixin Wu +9 · 5 citations
    Computer Science · #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Hate Speech and Cyberbullying Detection #Machine Learning (cs.LG)
  13. In ChatGPT We Trust? Measuring and Characterizing the Reliability of ChatGPT
    2023/04/18 by Xinyue Shen, Shen, Xinyue, Zeyuan Chen +5 · 2 citations
    Computer Science · Medicine · #Topic Modeling #Artificial Intelligence in Healthcare and Education #Explainable Artificial Intelligence (XAI)