Yaozhong Gan
- Stabilizing Q Learning Via Soft Mellowmax Operator
2020/12/17 by Yaozhong Gan, Gan, Yaozhong, Zhe Zhang +3 · 1 citation
Computer Science · #Adaptive Dynamic Programming Control #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
- Reflective Policy Optimization
2024/06/06 by Yaozhong Gan, Gan, Yaozhong, Renye Yan +5 · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #Distributed and Parallel Computing Systems #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- The Exploration-Exploitation Dilemma Revisited: An Entropy Perspective
2024/08/19 by Renye Yan, Yan, Renye, Yaozhong Gan +11 · 1 citation
Engineering · Physics and Astronomy · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Military Defense Systems Analysis #Opinion Dynamics and Social Influence