Daniel Kang
- LLM Agents can Autonomously Exploit One-day Vulnerabilities
2024/04/11 by Richard Fang, Rohan Bindu, Fang, Richard +5 · 9 voices · 24 citations
#cs.CR #cs.AI
- LLM Agents can Autonomously Hack Websites
2024/02/06 by Richard Fang, Rohan Bindu, Fang, Richard +7 · 8 voices · 23 citations
Business, Management and Accounting · Computer Science · #Blockchain Technology Applications and Security #Digital Rights Management and Security #FinTech, Crowdfunding, Digital Finance #cs.AI #cs.CR
- Measuring Agents in Production
2025/12/02 by Melissa Z. Pan, Pan, Melissa Z., Negar Arabzadeh +47 · 11 voices · 4 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computers and Society (cs.CY) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Software Engineering (cs.SE) #cs.AI #cs.CY #cs.LG #cs.SE
- InjecAgent: Benchmarking Indirect Prompt Injections in Tool-Integrated Large Language Model Agents
2024/03/05 by Qiusi Zhan, Zhan, Qiusi, Zhixiang Liang +5 · 76 citations
Computer Science · Business, Management and Accounting · #Topic Modeling #Natural Language Processing Techniques #Business Process Modeling and Analysis
- Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks
2023/02/11 by Daniel Kang, Xuechen Li, Kang, Daniel +9 · 40 citations
Computer Science · #Topic Modeling #Software Engineering Research #Natural Language Processing Techniques
- Holistic Agent Leaderboard: The Missing Infrastructure for AI Agent Evaluation
2025/10/13 by Sayash Kapoor, Benedikt Stroebl, Kapoor, Sayash +63 · 3 voices · 9 citations
Computer Science · #Multi-Agent Systems and Negotiation
- Teams of LLM Agents can Exploit Zero-Day Vulnerabilities
2024/06/02 by Yuxuan Zhu, Antony Kellermann, Zhu, Yuxuan +11 · 3 voices · 23 citations
Computer Science · #Blockchain Technology Applications and Security
- MLPerf Training Benchmark
2019/10/02 by Peter Mattson, Christine Cheng, Mattson, Peter +71 · 19 citations
Computer Science · #Advanced Neural Network Applications #Data Stream Mining Techniques #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Data Classification #Performance (cs.PF)
- A Safe Harbor for AI Evaluation and Red Teaming
2024/03/07 by Shayne Longpre, Sayash Kapoor, Longpre, Shayne +43 · 1 voice · 15 citations
Computer Science · #Explainable Artificial Intelligence (XAI)
- Adaptive Attacks Break Defenses Against Indirect Prompt Injection Attacks on LLM Agents
2025/02/27 by Qiusi Zhan, Zhan, Qiusi, Richard Fang +5 · 26 citations
Computer Science · #Blockchain Technology Applications and Security #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Network Security and Intrusion Detection #Security and Verification in Computing
- Scaling up Trustless DNN Inference with Zero-Knowledge Proofs
2022/10/17 by Daniel Kang, Tatsunori Hashimoto, Kang, Daniel +5 · 6 citations
Medicine · Computer Science · #COVID-19 diagnosis using AI #Medical Imaging Techniques and Applications #Advanced Neural Network Applications
- Improved Natural Language Generation via Loss Truncation
2020/04/30 by Daniel Kang, Kang, Daniel, Tatsunori Hashimoto +1 · 3 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Topic Modeling
- Network Offloading Policies for Cloud Robotics: a Learning-based\n Approach
2019/02/15 by Sandeep Chinchali, Chinchali, Sandeep, Apoorva Sharma +15 · 2 citations
Computer Science · Engineering · #IoT and Edge/Fog Computing #Robotics and Automated Systems #Advanced Neural Network Applications
- UTBoost: Rigorous Evaluation of Coding Agents on SWE-Bench
2025/06/10 by Bin Yu, Yu, Boxi, Pinjia He +4 · 6 citations
Computer Science · #Computation and Language (cs.CL) #D.0 #FOS: Computer and information sciences #I.2 #Software Engineering (cs.SE) #Software Engineering Research #Software Testing and Debugging Techniques #Topic Modeling
- NoScope: Optimizing Neural Network Queries over Video at Scale
2017/03/07 by Daniel Kang, Kang, Daniel, John Emmons +7 · 2 citations
Computer Science · #Advanced Neural Network Applications #Computer Vision and Pattern Recognition (cs.CV) #Databases (cs.DB) #FOS: Computer and information sciences #Human Pose and Action Recognition #Video Surveillance and Tracking Methods
- Trustless Audits without Revealing Data or Models
2024/04/06 by Suppakit Waiwitlikhit, Waiwitlikhit, Suppakit, Ion Stoica +7 · 3 citations
Computer Science · #Blockchain Technology Applications and Security
- ELT-Bench: An End-to-End Benchmark for Evaluating AI Agents on ELT Pipelines
2025/04/07 by Yuxuan Zhu, Jin, Tengjun, Daniel Kang +2 · 5 citations
Computer Science · Decision Sciences · #Advanced Database Systems and Queries #Artificial Intelligence (cs.AI) #Databases (cs.DB) #FOS: Computer and information sciences #Scientific Computing and Data Management #Semantic Web and Ontologies
- Testing Robustness Against Unforeseen Adversaries
2019/08/21 by Max Kaufmann, Kaufmann, Max, Daniel Kang +21 · 1 voice
Computer Science · Mathematics · #Computer Vision and Pattern Recognition (cs.CV) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #cs.CR #cs.CV #cs.LG #stat.ML
- REPRO-Bench: Can Agentic AI Systems Assess the Reproducibility of Social Science Research?
2025/07/25 by Chuxuan Hu, Liyun Zhang, Hu, Chuxuan +8 · 1 citation
Decision Sciences · Biochemistry, Genetics and Molecular Biology · Computer Science · #Scientific Computing and Data Management #Biomedical Text Mining and Ontologies #Research Data Management Practices
- Non-vacuous Generalization Bounds for Reinforcement Learning with Verifiable Rewards
2026/07/16 by Yuxuan Zhu, Rohan Alur, Daniel Kang
#cs.LG #cs.AI