vix.ing · top · new · best · stats · spec

Jin, Tiancheng

  1. Learning Adversarial MDPs with Bandit Feedback and Unknown Transition
    2019/12/03 by Chi Jin, Jin, Chi, Tiancheng Jin +7 · 11 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #I.2.6 #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Reinforcement Learning in Robotics
  2. The best of both worlds: stochastic and adversarial episodic MDPs with unknown transition
    2021/06/08 by Tiancheng Jin, Longbo Huang, Jin, Tiancheng +3 · 3 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #I.2.6 #Machine Learning (cs.LG) #Machine Learning and Algorithms
  3. Deep Reinforcement Learning for Multi-Driver Vehicle Dispatching and Repositioning Problem
    2019/11/25 by Holler, John, Vuorio, Risto, Qin, Zhiwei +6 · 2 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  4. Improved Best-of-Both-Worlds Guarantees for Multi-Armed Bandits: FTRL with General Regularizers and Multiple Optimal Arms
    2023/02/27 by Tiancheng Jin, Jin, Tiancheng, Junyan Liu +3 · 3 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #I.2.6 #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
  5. Simultaneously Learning Stochastic and Adversarial Episodic MDPs with Known Transition
    2020/06/10 by Jin, Tiancheng, Luo, Haipeng · 3 citations
    #FOS: Computer and information sciences #I.2.6 #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  6. No-Regret Online Reinforcement Learning with Adversarial Losses and Transitions
    2023/05/27 by Jin, Tiancheng, Liu, Junyan, Rouyer, Chloé +3 · 2 citations
    #FOS: Computer and information sciences #I.2.6 #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  7. Near-Optimal Regret for Adversarial MDP with Delayed Bandit Feedback
    2022/01/31 by Tiancheng Jin, Jin, Tiancheng, Tal Lancewicki +7 · 1 citation
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Algorithms #Reinforcement Learning in Robotics
  8. Heterogeneous Directed Hypergraph Neural Network over abstract syntax tree (AST) for Code Classification
    2023/05/07 by Yang, Guang, Jin, Tiancheng, Dou, Liang · 1 citation
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Software Engineering (cs.SE)