vix.ing · top · new · best · stats · spec

Tian Xu

  1. ReMax: A Simple, Effective, and Efficient Reinforcement Learning Method for Aligning Large Language Models
    2023/10/16 by Ziniu Li, Li, Ziniu, Tian Xu +10 · 44 citations
    Computer Science · #Topic Modeling #Natural Language Processing Techniques #Speech Recognition and Synthesis
  2. Preserving Diversity in Supervised Fine-Tuning of Large Language Models
    2024/08/29 by Ziniu Li, Congliang Chen, Li, Ziniu +11 · 15 citations
    Physics and Astronomy · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Magnetic confinement fusion research
  3. How Can Reinforcement Learning Achieve Expert-level Placement?
    2026/04/28 by Ruo-Tong Chen, Ke Xue, Chengrui Gao +7 · 3 voices
    #cs.AR #cs.AI #cs.LG
  4. On Value Discrepancy of Imitation Learning
    2019/11/16 by Tian Xu, Ziniu Li, Xu, Tian +3 · 2 citations
    Computer Science · #Reinforcement Learning in Robotics #Generative Adversarial Networks and Image Synthesis #Adversarial Robustness in Machine Learning