Guangchen Lan
- Asynchronous Federated Reinforcement Learning with Policy Gradient Updates: Algorithm Design and Convergence Analysis
2024/04/09 by Guangchen Lan, Dong-Jun Han, Lan, Guangchen +7 · 4 citations
Computer Science · Engineering · #Reinforcement Learning in Robotics #Advanced Memory and Neural Computing
- Bridging SFT and DPO for Diffusion Model Alignment with Self-Sampling Preference Optimization
2024/10/07 by Daoan Zhang, Zhang, Daoan, Guangchen Lan +19 · 2 citations
Social Sciences · #Computer Vision and Pattern Recognition (cs.CV) #EU Law and Policy Analysis #FOS: Computer and information sciences #I.2.10 #I.2.6 #I.4.0 #I.5.0 #Machine Learning (cs.LG)