Ruibin Xiong
- From Novice to Expert: LLM Agent Policy Optimization via Step-wise Reinforcement Learning
2024/11/06 by Zhirui Deng, Deng, Zhirui, Zhicheng Dou +11 · 6 citations
Computer Science · #Multi-Agent Systems and Negotiation
- ArtifactsBench: Bridging the Visual-Interactive Gap in LLM Code Generation Evaluation
2025/07/07 by Chenchen Zhang, Yuhang Li, Zhang, Chenchen +36 · 8 citations
Computer Science · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Software Engineering (cs.SE)
- When Does Group Invariant Learning Survive Spurious Correlations?
2022/06/29 by Yimeng Chen, Ruibin Xiong, Chen, Yimeng +5 · 1 citation
Computer Science · #Anomaly Detection Techniques and Applications #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Face and Expression Recognition #Machine Learning (cs.LG) #Machine Learning (stat.ML)