Nishant Balepur
- The Prompt Report: A Systematic Survey of Prompt Engineering Techniques
2024/06/06 by Sander Schulhoff, Schulhoff, Sander, Michael Ilie +62 · 16 voices · 61 citations
Computer Science · Medicine · #Artificial Intelligence in Healthcare and Education #Explainable Artificial Intelligence (XAI) #Topic Modeling #cs.AI #cs.CL
- Which of These Best Describes Multiple Choice Evaluation with LLMs? A) Forced B) Flawed C) Fixable D) All of the Above
2025/02/19 by Nishant Balepur, Balepur, Nishant, Rachel Rudinger +3 · 14 citations
Social Sciences · Economics, Econometrics and Finance · #Legal Education and Practice Innovations #Artificial Intelligence in Law #Occupational and Professional Licensing Regulation
- AstaBench: Rigorous Benchmarking of AI Agents with a Scientific Research Suite
2025/10/24 by Jonathan Bragg, Mike D'Arcy, Bragg, Jonathan +75 · 2 voices · 3 citations
#cs.AI #cs.CL
- Plausibly Problematic Questions in Multiple-Choice Benchmarks for Commonsense Reasoning
2024/10/06 by Shramay Palta, Palta, Shramay, Nishant Balepur +9 · 4 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Logic, Reasoning, and Knowledge
- Test-Time Reasoners Are Strategic Multiple-Choice Test-Takers
2025/10/09 by Nishant Balepur, Atrey Desai, Balepur, Nishant +3 · 2 voices
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #cs.CL
- Text Fact Transfer
2023/10/23 by Nishant Balepur, Balepur, Nishant, Jie Huang +3 · 1 citation
Computer Science · #Topic Modeling #Natural Language Processing Techniques #Text Readability and Simplification
- (Im)Paired Programming: Coding Agents Improve Productivity but Harm Understanding
2026/07/29 by Nishant Balepur, Connor Baumler, Valerie Chen +3
Computer Science · #cs.CL #cs.HC