Cui, Shiyao
- Agent-SafetyBench: Evaluating the Safety of LLM Agents
2024/12/19 by Zhexin Zhang, Shiyao Cui, Zhang, Zhexin +11 · 2 voices · 46 citations
Computer Science · Engineering · #Safety Systems Engineering in Autonomy #cs.CL
- NACL: A General and Effective KV Cache Eviction Framework for LLMs at Inference Time
2024/08/07 by Yilong Chen, Guoxia Wang, Chen, Yilong +17 · 12 citations
Computer Science · Social Sciences · Business, Management and Accounting · #Digital Rights Management and Security #Artificial Intelligence in Law #Financial Distress and Bankruptcy Prediction
- Human Decision-making is Susceptible to AI-driven Manipulation
2025/02/11 by Sabour, Sahand, Liu, June M., Liu, Siyang +13 · 6 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computers and Society (cs.CY) #FOS: Computer and information sciences #Human-Computer Interaction (cs.HC)
- How Should We Enhance the Safety of Large Reasoning Models: An Empirical Study
2025/05/21 by Zhang, Zhexin, Loye, Xian Qi, Huang, Victor Shea-Jay +8 · 9 citations
#Computation and Language (cs.CL) #FOS: Computer and information sciences
- AISafetyLab: A Comprehensive Framework for AI Safety Evaluation and Improvement
2025/02/24 by Zhang, Zhexin, Lei, Leqi, Yang, Junxiao +13 · 4 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
- Event Causality Extraction with Event Argument Correlations
2023/01/27 by Shiyao Cui, Jiawei Sheng, Cui, Shiyao +9 · 1 citation
Computer Science · #Advanced Text Analysis Techniques #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Topic Modeling
- Dual-Gated Fusion with Prefix-Tuning for Multi-Modal Relation Extraction
2023/06/19 by Qian Li, Shu Guo, Li, Qian +9 · 1 citation
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Topic Modeling
- ShieldVLM: Safeguarding the Multimodal Implicit Toxicity via Deliberative Reasoning with LVLMs
2025/05/20 by Shiyao Cui, Cui, Shiyao, Qinglin Zhang +14 · 3 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Multimedia (cs.MM) #Natural Language Processing Techniques
- LongSafety: Evaluating Long-Context Safety of Large Language Models
2025/02/24 by Lu, Yida, Cheng, Jiale, Zhang, Zhexin +7 · 2 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
- Guiding not Forcing: Enhancing the Transferability of Jailbreaking Attacks on LLMs via Removing Superfluous Constraints
2025/02/25 by Yang, Junxiao, Zhang, Zhexin, Cui, Shiyao +2 · 3 citations
#Artificial Intelligence (cs.AI) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- From Theft to Bomb-Making: The Ripple Effect of Unlearning in Defending Against Jailbreak Attacks
2024/07/03 by Zhang, Zhexin, Yang, Junxiao, Lu, Yida +5 · 1 citation
#Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- BARREL: Boundary-Aware Reasoning for Factual and Reliable LRMs
2025/05/18 by Junxiao Yang, Yang, Junxiao, Jiyuan Tu +20 · 2 citations
Computer Science · Decision Sciences · #Natural Language Processing Techniques #Topic Modeling #Data Quality and Management
- Global Challenge for Safe and Secure LLMs Track 1
2024/11/21 by Xiaojun Jia, Yihao Huang, Jia, Xiaojun +57 · 1 citation
Computer Science · #Law, AI, and Intellectual Property
- JPS: Jailbreak Multimodal Large Language Models with Collaborative Visual Perturbation and Textual Steering
2025/08/07 by Chen, Renmiao, Cui, Shiyao, Huang, Xuancheng +7 · 2 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #I.2.7 #K.4.1 #K.6.5 #Multimedia (cs.MM)