vix.ing · top · new · best · stats · spec

Daniel Kang

  1. LLM Agents can Autonomously Exploit One-day Vulnerabilities
    2024/04/11 by Richard Fang, Rohan Bindu, Fang, Richard +5 · 9 voices · 24 citations
    #cs.CR #cs.AI
  2. LLM Agents can Autonomously Hack Websites
    2024/02/06 by Richard Fang, Rohan Bindu, Fang, Richard +7 · 8 voices · 23 citations
    Business, Management and Accounting · Computer Science · #Blockchain Technology Applications and Security #Digital Rights Management and Security #FinTech, Crowdfunding, Digital Finance #cs.AI #cs.CR
  3. Measuring Agents in Production
    2025/12/02 by Melissa Z. Pan, Pan, Melissa Z., Negar Arabzadeh +47 · 11 voices · 4 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computers and Society (cs.CY) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Software Engineering (cs.SE) #cs.AI #cs.CY #cs.LG #cs.SE
  4. InjecAgent: Benchmarking Indirect Prompt Injections in Tool-Integrated Large Language Model Agents
    2024/03/05 by Qiusi Zhan, Zhan, Qiusi, Zhixiang Liang +5 · 76 citations
    Computer Science · Business, Management and Accounting · #Topic Modeling #Natural Language Processing Techniques #Business Process Modeling and Analysis
  5. Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks
    2023/02/11 by Daniel Kang, Xuechen Li, Kang, Daniel +9 · 40 citations
    Computer Science · #Topic Modeling #Software Engineering Research #Natural Language Processing Techniques
  6. Holistic Agent Leaderboard: The Missing Infrastructure for AI Agent Evaluation
    2025/10/13 by Sayash Kapoor, Benedikt Stroebl, Kapoor, Sayash +63 · 3 voices · 9 citations
    Computer Science · #Multi-Agent Systems and Negotiation
  7. Teams of LLM Agents can Exploit Zero-Day Vulnerabilities
    2024/06/02 by Yuxuan Zhu, Antony Kellermann, Zhu, Yuxuan +11 · 3 voices · 23 citations
    Computer Science · #Blockchain Technology Applications and Security
  8. MLPerf Training Benchmark
    2019/10/02 by Peter Mattson, Christine Cheng, Mattson, Peter +71 · 19 citations
    Computer Science · #Advanced Neural Network Applications #Data Stream Mining Techniques #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Data Classification #Performance (cs.PF)
  9. A Safe Harbor for AI Evaluation and Red Teaming
    2024/03/07 by Shayne Longpre, Sayash Kapoor, Longpre, Shayne +43 · 1 voice · 15 citations
    Computer Science · #Explainable Artificial Intelligence (XAI)
  10. Adaptive Attacks Break Defenses Against Indirect Prompt Injection Attacks on LLM Agents
    2025/02/27 by Qiusi Zhan, Zhan, Qiusi, Richard Fang +5 · 26 citations
    Computer Science · #Blockchain Technology Applications and Security #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Network Security and Intrusion Detection #Security and Verification in Computing
  11. Scaling up Trustless DNN Inference with Zero-Knowledge Proofs
    2022/10/17 by Daniel Kang, Tatsunori Hashimoto, Kang, Daniel +5 · 6 citations
    Medicine · Computer Science · #COVID-19 diagnosis using AI #Medical Imaging Techniques and Applications #Advanced Neural Network Applications
  12. Improved Natural Language Generation via Loss Truncation
    2020/04/30 by Daniel Kang, Kang, Daniel, Tatsunori Hashimoto +1 · 3 citations
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Topic Modeling
  13. Network Offloading Policies for Cloud Robotics: a Learning-based\n Approach
    2019/02/15 by Sandeep Chinchali, Chinchali, Sandeep, Apoorva Sharma +15 · 2 citations
    Computer Science · Engineering · #IoT and Edge/Fog Computing #Robotics and Automated Systems #Advanced Neural Network Applications
  14. UTBoost: Rigorous Evaluation of Coding Agents on SWE-Bench
    2025/06/10 by Bin Yu, Yu, Boxi, Pinjia He +4 · 6 citations
    Computer Science · #Computation and Language (cs.CL) #D.0 #FOS: Computer and information sciences #I.2 #Software Engineering (cs.SE) #Software Engineering Research #Software Testing and Debugging Techniques #Topic Modeling
  15. NoScope: Optimizing Neural Network Queries over Video at Scale
    2017/03/07 by Daniel Kang, Kang, Daniel, John Emmons +7 · 2 citations
    Computer Science · #Advanced Neural Network Applications #Computer Vision and Pattern Recognition (cs.CV) #Databases (cs.DB) #FOS: Computer and information sciences #Human Pose and Action Recognition #Video Surveillance and Tracking Methods
  16. Trustless Audits without Revealing Data or Models
    2024/04/06 by Suppakit Waiwitlikhit, Waiwitlikhit, Suppakit, Ion Stoica +7 · 3 citations
    Computer Science · #Blockchain Technology Applications and Security
  17. ELT-Bench: An End-to-End Benchmark for Evaluating AI Agents on ELT Pipelines
    2025/04/07 by Yuxuan Zhu, Jin, Tengjun, Daniel Kang +2 · 5 citations
    Computer Science · Decision Sciences · #Advanced Database Systems and Queries #Artificial Intelligence (cs.AI) #Databases (cs.DB) #FOS: Computer and information sciences #Scientific Computing and Data Management #Semantic Web and Ontologies
  18. Testing Robustness Against Unforeseen Adversaries
    2019/08/21 by Max Kaufmann, Kaufmann, Max, Daniel Kang +21 · 1 voice
    Computer Science · Mathematics · #Computer Vision and Pattern Recognition (cs.CV) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #cs.CR #cs.CV #cs.LG #stat.ML
  19. REPRO-Bench: Can Agentic AI Systems Assess the Reproducibility of Social Science Research?
    2025/07/25 by Chuxuan Hu, Liyun Zhang, Hu, Chuxuan +8 · 1 citation
    Decision Sciences · Biochemistry, Genetics and Molecular Biology · Computer Science · #Scientific Computing and Data Management #Biomedical Text Mining and Ontologies #Research Data Management Practices
  20. Non-vacuous Generalization Bounds for Reinforcement Learning with Verifiable Rewards
    2026/07/16 by Yuxuan Zhu, Rohan Alur, Daniel Kang
    #cs.LG #cs.AI