vix.ing · top · new · best · stats · spec

Yunchang Yang

  1. Nearly Optimal Policy Optimization with Stable at Any Time Guarantee
    2021/12/21 by Tianhao Wu, Yunchang Yang, Wu, Tianhao +9 · 1 citation
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics