Cuadron, Alejandro
- JudgeBench: A Benchmark for Evaluating LLM-based Judges
2024/10/16 by Siyuan Zhuang, Tan, Sijun, K. Leon Montgomery +11 · 62 citations
Social Sciences · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Law #Computation and Language (cs.CL) #FOS: Computer and information sciences #Legal Education and Practice Innovations #Legal Systems and Judicial Processes #Machine Learning (cs.LG)
- The Danger of Overthinking: Examining the Reasoning-Action Dilemma in Agentic Tasks
2025/02/12 by Alejandro Cuadron, Dacheng Li, Cuadron, Alejandro +28 · 36 citations
Social Sciences · #Experimental Behavioral Economics Studies
- HashAttention: Semantic Sparsity for Faster Inference
2024/12/19 by Aditya Desai, Desai, Aditya, Shuo Yang +9 · 5 citations
Computer Science · Medicine · #Cryptography and Data Security #Big Data and Digital Economy #COVID-19 diagnosis using AI
- vCache: Verified Semantic Prompt Caching
2025/02/06 by Luis Gaspar Schroeder, Schroeder, Luis Gaspar, Aditya Desai +18 · 1 citation
Computer Science · #Caching and Content Delivery #Service-Oriented Architecture and Web Services #Cognitive Computing and Networks
- vAttention: Verified Sparse Attention
2025/10/07 by Aditya Desai, Kumar Krishna Agrawal, Desai, Aditya +13 · 2 citations
Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- SABER: Small Actions, Big Errors -- Safeguarding Mutating Steps in LLM Agents
2025/11/26 by Cuadron, Alejandro, Yu, Pengfei, Liu, Yang +1 · 1 citation
Computer Science · Social Sciences · #Artificial Intelligence (cs.AI) #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multi-Agent Systems and Negotiation