Jin, Tiancheng
- Learning Adversarial MDPs with Bandit Feedback and Unknown Transition
2019/12/03 by Chi Jin, Jin, Chi, Tiancheng Jin +7 · 11 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #I.2.6 #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Reinforcement Learning in Robotics
- The best of both worlds: stochastic and adversarial episodic MDPs with unknown transition
2021/06/08 by Tiancheng Jin, Longbo Huang, Jin, Tiancheng +3 · 3 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #I.2.6 #Machine Learning (cs.LG) #Machine Learning and Algorithms
- Deep Reinforcement Learning for Multi-Driver Vehicle Dispatching and Repositioning Problem
2019/11/25 by Holler, John, Vuorio, Risto, Qin, Zhiwei +6 · 2 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Improved Best-of-Both-Worlds Guarantees for Multi-Armed Bandits: FTRL with General Regularizers and Multiple Optimal Arms
2023/02/27 by Tiancheng Jin, Jin, Tiancheng, Junyan Liu +3 · 3 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #I.2.6 #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
- Simultaneously Learning Stochastic and Adversarial Episodic MDPs with Known Transition
2020/06/10 by Jin, Tiancheng, Luo, Haipeng · 3 citations
#FOS: Computer and information sciences #I.2.6 #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- No-Regret Online Reinforcement Learning with Adversarial Losses and Transitions
2023/05/27 by Jin, Tiancheng, Liu, Junyan, Rouyer, Chloé +3 · 2 citations
#FOS: Computer and information sciences #I.2.6 #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Near-Optimal Regret for Adversarial MDP with Delayed Bandit Feedback
2022/01/31 by Tiancheng Jin, Jin, Tiancheng, Tal Lancewicki +7 · 1 citation
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Algorithms #Reinforcement Learning in Robotics
- Heterogeneous Directed Hypergraph Neural Network over abstract syntax tree (AST) for Code Classification
2023/05/07 by Yang, Guang, Jin, Tiancheng, Dou, Liang · 1 citation
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Software Engineering (cs.SE)