vix.ing · top · new · best · stats · spec

Tiancheng Jin

  1. Learning Adversarial MDPs with Bandit Feedback and Unknown Transition
    2019/12/03 by Chi Jin, Jin, Chi, Tiancheng Jin +7 · 6 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #I.2.6 #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Reinforcement Learning in Robotics
  2. The best of both worlds: stochastic and adversarial episodic MDPs with unknown transition
    2021/06/08 by Tiancheng Jin, Jin, Tiancheng, Longbo Huang +3 · 2 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #I.2.6 #Machine Learning (cs.LG) #Machine Learning and Algorithms
  3. Near-Optimal Regret for Adversarial MDP with Delayed Bandit Feedback
    2022/01/31 by Tiancheng Jin, Tal Lancewicki, Jin, Tiancheng +7 · 1 citation
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Algorithms #Reinforcement Learning in Robotics