vix.ing · top · new · best · stats · spec

Cuadron, Alejandro

  1. JudgeBench: A Benchmark for Evaluating LLM-based Judges
    2024/10/16 by Siyuan Zhuang, Tan, Sijun, K. Leon Montgomery +11 · 62 citations
    Social Sciences · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Law #Computation and Language (cs.CL) #FOS: Computer and information sciences #Legal Education and Practice Innovations #Legal Systems and Judicial Processes #Machine Learning (cs.LG)
  2. The Danger of Overthinking: Examining the Reasoning-Action Dilemma in Agentic Tasks
    2025/02/12 by Alejandro Cuadron, Dacheng Li, Cuadron, Alejandro +28 · 36 citations
    Social Sciences · #Experimental Behavioral Economics Studies
  3. HashAttention: Semantic Sparsity for Faster Inference
    2024/12/19 by Aditya Desai, Desai, Aditya, Shuo Yang +9 · 5 citations
    Computer Science · Medicine · #Cryptography and Data Security #Big Data and Digital Economy #COVID-19 diagnosis using AI
  4. vCache: Verified Semantic Prompt Caching
    2025/02/06 by Luis Gaspar Schroeder, Schroeder, Luis Gaspar, Aditya Desai +18 · 1 citation
    Computer Science · #Caching and Content Delivery #Service-Oriented Architecture and Web Services #Cognitive Computing and Networks
  5. vAttention: Verified Sparse Attention
    2025/10/07 by Aditya Desai, Kumar Krishna Agrawal, Desai, Aditya +13 · 2 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  6. SABER: Small Actions, Big Errors -- Safeguarding Mutating Steps in LLM Agents
    2025/11/26 by Cuadron, Alejandro, Yu, Pengfei, Liu, Yang +1 · 1 citation
    Computer Science · Social Sciences · #Artificial Intelligence (cs.AI) #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multi-Agent Systems and Negotiation