Tian Xu
- ReMax: A Simple, Effective, and Efficient Reinforcement Learning Method for Aligning Large Language Models
2023/10/16 by Ziniu Li, Li, Ziniu, Tian Xu +10 · 44 citations
Computer Science · #Topic Modeling #Natural Language Processing Techniques #Speech Recognition and Synthesis
- Preserving Diversity in Supervised Fine-Tuning of Large Language Models
2024/08/29 by Ziniu Li, Congliang Chen, Li, Ziniu +11 · 15 citations
Physics and Astronomy · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Magnetic confinement fusion research
- How Can Reinforcement Learning Achieve Expert-level Placement?
2026/04/28 by Ruo-Tong Chen, Ke Xue, Chengrui Gao +7 · 3 voices
#cs.AR #cs.AI #cs.LG
- On Value Discrepancy of Imitation Learning
2019/11/16 by Tian Xu, Ziniu Li, Xu, Tian +3 · 2 citations
Computer Science · #Reinforcement Learning in Robotics #Generative Adversarial Networks and Image Synthesis #Adversarial Robustness in Machine Learning