Yunchang Yang
- Nearly Optimal Policy Optimization with Stable at Any Time Guarantee
2021/12/21 by Tianhao Wu, Yunchang Yang, Wu, Tianhao +9 · 1 citation
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics