Shibo Hong
- Toward Generalizable Evaluation in the LLM Era: A Survey Beyond Benchmarks
2025/04/26 by Yixin Cao, Shibo Hong, Cao, Yixin +49 · 13 citations
Computer Science · Social Sciences · #Computation and Language (cs.CL) #Computational and Text Analysis Methods #FOS: Computer and information sciences #Text Readability and Simplification #Topic Modeling
- Two Minds Better Than One: Collaborative Reward Modeling for LLM Alignment
2025/05/15 by Jiazheng Zhang, Zhang, Jiazheng, Zizhuo Zhang +20 · 4 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Recommender Systems and Techniques #Topic Modeling
- ICAE-Bench: Evaluating Coding Agents as Interactive Project Builders
2026/07/23 by Zhongyuan Peng, Dan Huang, Chuyu Zhang +8
#cs.AI