Boyi Wei
- Scaling Latent Reasoning via Looped Language Models
2025/10/29 by Rui-Jie Zhu, Zixuan Wang, Zhu, Rui-Jie +63 · 14 voices · 21 citations
#cs.CL
- Humanity's Last Exam
2025/01/24 by Long Phan, Alice Gatti, Phan, Long +2240 · 9 voices · 103 citations
#cs.LG #cs.AI #cs.CL
- SORRY-Bench: Systematically Evaluating Large Language Model Safety Refusal
2024/06/20 by Tinghao Xie, Xiangyu Qi, Xie, Tinghao +29 · 39 citations
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Natural Language Processing Techniques #Software Reliability and Analysis Research #Topic Modeling
- Holistic Agent Leaderboard: The Missing Infrastructure for AI Agent Evaluation
2025/10/13 by Sayash Kapoor, Benedikt Stroebl, Kapoor, Sayash +63 · 3 voices · 9 citations
Computer Science · #Multi-Agent Systems and Negotiation
- Assessing the Brittleness of Safety Alignment via Pruning and Low-Rank Modifications
2024/02/07 by Boyi Wei, Kaixuan Huang, Wei, Boyi +15 · 33 citations
Decision Sciences · Engineering · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Fatigue and fracture mechanics #Machine Learning (cs.LG) #Risk and Safety Analysis
- An Adversarial Perspective on Machine Unlearning for AI Safety
2024/09/26 by Jakub Łucki, Łucki, Jakub, Boyi Wei +9 · 3 voices · 19 citations
Computer Science · #Adversarial Robustness in Machine Learning #cs.AI #cs.CL #cs.CR #cs.LG
- Evaluating Copyright Takedown Methods for Language Models
2024/06/26 by Boyi Wei, Weijia Shi, Wei, Boyi +13 · 8 citations
Computer Science · #Computation and Language (cs.CL) #Digital Rights Management and Security #FOS: Computer and information sciences #Machine Learning (cs.LG)
- On Evaluating the Durability of Safeguards for Open-Weight LLMs
2024/12/10 by Xiangyu Qi, Boyi Wei, Qi, Xiangyu +17 · 2 voices · 7 citations
Computer Science · #Artificial Intelligence (cs.AI) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #cs.AI #cs.CR
- Dynamic Risk Assessments for Offensive Cybersecurity Agents
2025/05/23 by Boyi Wei, Wei, Boyi, Benedikt Stroebl +9 · 1 voice · 3 citations
Computer Science · Engineering · #Artificial Intelligence (cs.AI) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Information and Cyber Security #Network Security and Intrusion Detection #Smart Grid Security and Resilience #cs.AI #cs.CR
- IDO-VFI: Identifying Dynamics via Optical Flow Guidance for Video Frame Interpolation with Events
2023/05/17 by Chenyang Shi, Hanxiao Liu, Shi, Chenyang +11 · 1 citation
Computer Science · #Advanced Image Processing Techniques #Advanced Vision and Imaging #Anomaly Detection Techniques and Applications
- AI Risk Management Should Incorporate Both Safety and Security
2024/05/29 by Xiangyu Qi, Yangsibo Huang, Qi, Xiangyu +47 · 1 citation
Medicine · Social Sciences · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Healthcare and Education #Cryptography and Security (cs.CR) #Ethics and Social Impacts of AI #FOS: Computer and information sciences