Jiafan He
- Logarithmic Regret for Reinforcement Learning with Linear Function\n Approximation
2020/11/23 by Jiafan He, Dongruo Zhou, He, Jiafan +3 · 8 citations
Computer Science · Decision Sciences · Engineering · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Optimization and Control (math.OC) #Reinforcement Learning in Robotics #Smart Grid Energy Management
- Reinforcement Learning from Human Feedback with Active Queries
2024/02/14 by Kaixuan Ji, Ji, Kaixuan, Jiafan He +3 · 8 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Data Stream Mining Techniques #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Neural Networks and Applications #Optimization and Control (math.OC) #Reinforcement Learning in Robotics
- Variance-Dependent Regret Bounds for Linear Bandits and Reinforcement Learning: Adaptivity and Computational Efficiency
2023/02/21 by Heyang Zhao, Zhao, Heyang, Jiafan He +7 · 7 citations
Computer Science · Decision Sciences · Engineering · #Advanced Bandit Algorithms Research #Age of Information Optimization #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Optimization and Control (math.OC) #Smart Grid Energy Management
- Nearly Minimax Optimal Reinforcement Learning for Linear Markov Decision Processes
2022/12/12 by Jiafan He, Heyang Zhao, He, Jiafan +5 · 3 citations
Computer Science · Mathematics · Medicine · #Advanced Causal Inference Techniques #FOS: Computer and information sciences #FOS: Mathematics #Influenza Virus Research Studies #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Optimization and Control (math.OC) #Reinforcement Learning in Robotics
- Nearly Optimal Algorithms for Linear Contextual Bandits with Adversarial Corruptions
2022/05/13 by Jiafan He, Dongruo Zhou, He, Jiafan +5 · 2 citations
Decision Sciences · Computer Science · #Advanced Bandit Algorithms Research #Optimization and Search Problems #Stochastic Gradient Optimization Techniques
- Towards Robust Model-Based Reinforcement Learning Against Adversarial Corruption
2024/02/14 by Chenlu Ye, Ye, Chenlu, Jiafan He +5 · 3 citations
Computer Science · #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- A Nearly Optimal and Low-Switching Algorithm for Reinforcement Learning with General Function Approximation
2023/11/26 by Heyang Zhao, Jiafan He, Zhao, Heyang +3 · 2 citations
Computer Science · Decision Sciences · #Reinforcement Learning in Robotics #Advanced Bandit Algorithms Research #Optimization and Search Problems
- Nemotron-Cascade 2: Post-Training LLMs with Cascade RL and Multi-Domain On-Policy Distillation
2026/03/19 by Zhuolin Yang, Zihan Liu, Yang Chen +14 · 3 voices
Computer Science · #cs.CL #cs.AI #cs.LG
- Optimal Online Generalized Linear Regression with Stochastic Noise and Its Application to Heteroscedastic Bandits
2022/02/28 by Heyang Zhao, Zhao, Heyang, Dongruo Zhou +5 · 1 citation
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Distributed Sensor Networks and Detection Algorithms #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Optimization and Control (math.OC)
- Pessimistic Nonlinear Least-Squares Value Iteration for Offline Reinforcement Learning
2023/10/02 by Qiwei Di, Heyang Zhao, Di, Qiwei +5 · 1 citation
Decision Sciences · Mathematics · Computer Science · #Advanced Bandit Algorithms Research #COVID-19 epidemiological studies #Reinforcement Learning in Robotics
- Uniform-PAC Bounds for Reinforcement Learning with Linear Function\n Approximation
2021/06/22 by Jiafan He, Dongruo Zhou, He, Jiafan +3 · 1 citation
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Optimization and Control (math.OC) #Reinforcement Learning in Robotics
- Provably Efficient Representation Selection in Low-rank Markov Decision Processes: From Online to Offline RL
2021/06/22 by Weitong Zhang, Zhang, Weitong, Jiafan He +7 · 3 citations
Computer Science · Decision Sciences · Engineering · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Optimization and Control (math.OC) #Reinforcement Learning in Robotics #Smart Grid Energy Management