vix.ing · top · new · best · stats · spec

Cui, Shiyao

  1. Agent-SafetyBench: Evaluating the Safety of LLM Agents
    2024/12/19 by Zhexin Zhang, Shiyao Cui, Zhang, Zhexin +11 · 2 voices · 46 citations
    Computer Science · Engineering · #Safety Systems Engineering in Autonomy #cs.CL
  2. NACL: A General and Effective KV Cache Eviction Framework for LLMs at Inference Time
    2024/08/07 by Yilong Chen, Guoxia Wang, Chen, Yilong +17 · 12 citations
    Computer Science · Social Sciences · Business, Management and Accounting · #Digital Rights Management and Security #Artificial Intelligence in Law #Financial Distress and Bankruptcy Prediction
  3. Human Decision-making is Susceptible to AI-driven Manipulation
    2025/02/11 by Sabour, Sahand, Liu, June M., Liu, Siyang +13 · 6 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computers and Society (cs.CY) #FOS: Computer and information sciences #Human-Computer Interaction (cs.HC)
  4. How Should We Enhance the Safety of Large Reasoning Models: An Empirical Study
    2025/05/21 by Zhang, Zhexin, Loye, Xian Qi, Huang, Victor Shea-Jay +8 · 9 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  5. AISafetyLab: A Comprehensive Framework for AI Safety Evaluation and Improvement
    2025/02/24 by Zhang, Zhexin, Lei, Leqi, Yang, Junxiao +13 · 4 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
  6. Event Causality Extraction with Event Argument Correlations
    2023/01/27 by Shiyao Cui, Jiawei Sheng, Cui, Shiyao +9 · 1 citation
    Computer Science · #Advanced Text Analysis Techniques #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Topic Modeling
  7. Dual-Gated Fusion with Prefix-Tuning for Multi-Modal Relation Extraction
    2023/06/19 by Qian Li, Shu Guo, Li, Qian +9 · 1 citation
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Topic Modeling
  8. ShieldVLM: Safeguarding the Multimodal Implicit Toxicity via Deliberative Reasoning with LVLMs
    2025/05/20 by Shiyao Cui, Cui, Shiyao, Qinglin Zhang +14 · 3 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Multimedia (cs.MM) #Natural Language Processing Techniques
  9. LongSafety: Evaluating Long-Context Safety of Large Language Models
    2025/02/24 by Lu, Yida, Cheng, Jiale, Zhang, Zhexin +7 · 2 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
  10. Guiding not Forcing: Enhancing the Transferability of Jailbreaking Attacks on LLMs via Removing Superfluous Constraints
    2025/02/25 by Yang, Junxiao, Zhang, Zhexin, Cui, Shiyao +2 · 3 citations
    #Artificial Intelligence (cs.AI) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  11. From Theft to Bomb-Making: The Ripple Effect of Unlearning in Defending Against Jailbreak Attacks
    2024/07/03 by Zhang, Zhexin, Yang, Junxiao, Lu, Yida +5 · 1 citation
    #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  12. BARREL: Boundary-Aware Reasoning for Factual and Reliable LRMs
    2025/05/18 by Junxiao Yang, Yang, Junxiao, Jiyuan Tu +20 · 2 citations
    Computer Science · Decision Sciences · #Natural Language Processing Techniques #Topic Modeling #Data Quality and Management
  13. Global Challenge for Safe and Secure LLMs Track 1
    2024/11/21 by Xiaojun Jia, Yihao Huang, Jia, Xiaojun +57 · 1 citation
    Computer Science · #Law, AI, and Intellectual Property
  14. JPS: Jailbreak Multimodal Large Language Models with Collaborative Visual Perturbation and Textual Steering
    2025/08/07 by Chen, Renmiao, Cui, Shiyao, Huang, Xuancheng +7 · 2 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #I.2.7 #K.4.1 #K.6.5 #Multimedia (cs.MM)