Shang, Xuedong
- Gamification of Pure Exploration for Linear Bandits
2020/07/02 by Rémy Degenne, Degenne, Rémy, Pierre Ménard +5 · 8 citations
Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- UCB Momentum Q-learning: Correcting the bias without forgetting
2021/03/01 by Menard, Pierre, Domingues, Omar Darwiche, Shang, Xuedong +1 · 3 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Fixed-Confidence Guarantees for Bayesian Best-Arm Identification
2019/10/24 by Shang, Xuedong, de Heide, Rianne, Kaufmann, Emilie +2 · 2 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)