vix.ing · top · new · best · stats · spec

Nishant Balepur

  1. The Prompt Report: A Systematic Survey of Prompt Engineering Techniques
    2024/06/06 by Sander Schulhoff, Schulhoff, Sander, Michael Ilie +62 · 16 voices · 61 citations
    Computer Science · Medicine · #Artificial Intelligence in Healthcare and Education #Explainable Artificial Intelligence (XAI) #Topic Modeling #cs.AI #cs.CL
  2. Which of These Best Describes Multiple Choice Evaluation with LLMs? A) Forced B) Flawed C) Fixable D) All of the Above
    2025/02/19 by Nishant Balepur, Balepur, Nishant, Rachel Rudinger +3 · 14 citations
    Social Sciences · Economics, Econometrics and Finance · #Legal Education and Practice Innovations #Artificial Intelligence in Law #Occupational and Professional Licensing Regulation
  3. AstaBench: Rigorous Benchmarking of AI Agents with a Scientific Research Suite
    2025/10/24 by Jonathan Bragg, Mike D'Arcy, Bragg, Jonathan +75 · 2 voices · 3 citations
    #cs.AI #cs.CL
  4. Plausibly Problematic Questions in Multiple-Choice Benchmarks for Commonsense Reasoning
    2024/10/06 by Shramay Palta, Palta, Shramay, Nishant Balepur +9 · 4 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Logic, Reasoning, and Knowledge
  5. Test-Time Reasoners Are Strategic Multiple-Choice Test-Takers
    2025/10/09 by Nishant Balepur, Atrey Desai, Balepur, Nishant +3 · 2 voices
    Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #cs.CL
  6. Text Fact Transfer
    2023/10/23 by Nishant Balepur, Balepur, Nishant, Jie Huang +3 · 1 citation
    Computer Science · #Topic Modeling #Natural Language Processing Techniques #Text Readability and Simplification
  7. (Im)Paired Programming: Coding Agents Improve Productivity but Harm Understanding
    2026/07/29 by Nishant Balepur, Connor Baumler, Valerie Chen +3
    Computer Science · #cs.CL #cs.HC