Li, Dongbai
- RealSafe-R1: Safety-Aligned DeepSeek-R1 without Compromising Reasoning Capability
2025/04/14 by Yichi Zhang, Zhang, Yichi, Zihao Zeng +9 · 27 citations
Computer Science · #AI-based Problem Solving and Planning #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Semantic Web and Ontologies #Topic Modeling
- Sample Weight Averaging for Stable Prediction
2025/02/11 by Yu, Han, He, Yue, Xu, Renzhe +4 · 2 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG)
- The Emperor's New Clothes in Benchmarking? A Rigorous Examination of Mitigation Strategies for LLM Benchmark Data Contamination
2025/03/20 by Yifan Sun, Sun, Yifan, Han Wang +7 · 2 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Digital Rights Management and Security #FOS: Computer and information sciences #Machine Learning (cs.LG)