Gan, Yang
- GRAM: A Generative Foundation Reward Model for Reward Generalization
2025/06/17 by Chenglong Wang, Wang, Chenglong, Gan Yang +19 · 14 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Statistical and Computational Modeling
- RoVRM: A Robust Visual Reward Model Optimized via Auxiliary Textual Preference Data
2024/08/22 by Wang, Chenglong, Gan, Yang, Huo, Yifu +9 · 3 citations
#Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
- LRHP: Learning Representations for Human Preferences via Preference Pairs
2024/10/06 by Chenglong Wang, Gan Yang, Wang, Chenglong +17 · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #Bayesian Modeling and Causal Inference #Computation and Language (cs.CL) #Data Management and Algorithms #FOS: Computer and information sciences