Xiong, Zidi
- DecodingTrust: A Comprehensive Assessment of Trustworthiness in GPT Models
2023/06/20 by Boxin Wang, Wang, Boxin, Weixin Chen +35 · 58 citations
Computer Science · Medicine · Social Sciences · #Adversarial Robustness in Machine Learning #Artificial Intelligence in Healthcare and Education #Ethics and Social Impacts of AI
- GuardAgent: Safeguard LLM Agents by a Guard Agent via Knowledge-Enabled Reasoning
2024/06/13 by Zhen Xiang, Xiang, Zhen, Linzhi Zheng +21 · 36 citations
Computer Science · Engineering · Social Sciences · #Access Control and Trust #FOS: Computer and information sciences #Machine Learning (cs.LG) #Network Security and Intrusion Detection #Smart Grid Security and Resilience
- BadChain: Backdoor Chain-of-Thought Prompting for Large Language Models
2024/01/20 by Zhen Xiang, Xiang, Zhen, Fengqing Jiang +9 · 15 citations
Computer Science · Medicine · #Adversarial Robustness in Machine Learning #Artificial Intelligence in Healthcare and Education #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Topic Modeling
- RigorLLM: Resilient Guardrails for Large Language Models against Undesired Content
2024/03/19 by Yuan, Zhuowen, Xiong, Zidi, Zeng, Yi +4 · 13 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- How Memory Management Impacts LLM Agents: An Empirical Study of Experience-Following Behavior
2025/05/21 by Xiong, Zidi, Lin, Yuping, Xie, Wenya +5 · 16 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences
- When Models Reason in Your Language: Controlling Thinking Language Comes at the Cost of Accuracy
2025/05/28 by Jirui Qi, Qi, Jirui, Shan Chen +9 · 1 voice · 10 citations
Computer Science · #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Topic Modeling #cs.CL
- Measuring the Faithfulness of Thinking Drafts in Large Reasoning Models
2025/05/19 by Zidi Xiong, Xiong, Zidi, Shan Chen +5 · 6 citations
Computer Science · #Artificial Intelligence (cs.AI) #Constraint Satisfaction and Optimization #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Topic Modeling
- MMDT: Decoding the Trustworthiness and Safety of Multimodal Foundation Models
2025/03/19 by Chejian Xu, Jiawei Zhang, Xu, Chejian +44 · 3 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Multi-Agent Systems and Negotiation
- CBD: A Certified Backdoor Detector Based on Local Dominant Probability
2023/10/26 by Xiang, Zhen, Xiong, Zidi, Li, Bo · 1 citation
#Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)