Runlong Zhou
- Preference-Based Multi-Agent Reinforcement Learning: Data Coverage and Algorithmic Techniques
2024/09/01 by N Zhang, Zhang, Natalia, Xinqi Wang +9 · 3 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computer Science and Game Theory (cs.GT) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multiagent Systems (cs.MA) #Reinforcement Learning in Robotics
- Transformers are Efficient Compilers, Provably
2024/10/07 by Xiyu Zhai, Runlong Zhou, Zhai, Xiyu +5 · 1 voice
#cs.PL #cs.LG
- Reflect-RL: Two-Player Online RL Fine-Tuning for LMs
2024/02/20 by Runlong Zhou, Simon S. Du, Zhou, Runlong +3 · 2 citations
Engineering · #Computation and Language (cs.CL) #Experimental Learning in Engineering #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Understanding the Performance Gap in Preference Learning: A Dichotomy of RLHF and DPO
2025/05/26 by Ruizhe Shi, Shi, Ruizhe, Song, Minhak +8 · 2 citations
Decision Sciences · Computer Science · #Multi-Criteria Decision Making #Semantic Web and Ontologies
- Asymptotically Optimal Regret for Reinforcement Learning without Horizon Dependence
2026/07/22 by Runlong Zhou, Zihan Zhang, Maryam Fazel +1
#cs.LG #stat.ML