Duan, Yawen
- Adversarial Policies Beat Superhuman Go AIs
2022/11/01 by Tony Tong Wang, Tony T. Wang, Adam Gleave +21 · 18 voices · 7 citations
Computer Science · Social Sciences · #Adversarial Robustness in Machine Learning #Advanced Malware Detection Techniques #Ethics and Social Impacts of AI
- AI Alignment: A Comprehensive Survey
2023/10/30 by Ji, Jiaming, Qiu, Tianyi, Chen, Boyuan +23 · 39 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences
- TransNAS-Bench-101: Improving Transferability and Generalizability of Cross-Task Neural Architecture Search
2021/05/25 by Duan, Yawen, Chen, Xin, Xu, Hang +4 · 3 citations
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- On The Fragility of Learned Reward Functions
2023/01/09 by McKinney, Lev, Duan, Yawen, Krueger, David +1 · 2 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Bare Minimum Mitigations for Autonomous AI Development
2025/04/21 by Clymer, Joshua, Duan, Isabella, Cundy, Chris +10 · 1 citation
#Computers and Society (cs.CY) #FOS: Computer and information sciences
- Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report
2025/07/22 by Lab, Shanghai AI, :, Chen, Xiaoyang +35 · 4 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG)