vix.ing · top · new · best · stats · spec

Jiafan He

  1. Logarithmic Regret for Reinforcement Learning with Linear Function\n Approximation
    2020/11/23 by Jiafan He, Dongruo Zhou, He, Jiafan +3 · 8 citations
    Computer Science · Decision Sciences · Engineering · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Optimization and Control (math.OC) #Reinforcement Learning in Robotics #Smart Grid Energy Management
  2. Reinforcement Learning from Human Feedback with Active Queries
    2024/02/14 by Kaixuan Ji, Ji, Kaixuan, Jiafan He +3 · 8 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Data Stream Mining Techniques #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Neural Networks and Applications #Optimization and Control (math.OC) #Reinforcement Learning in Robotics
  3. Variance-Dependent Regret Bounds for Linear Bandits and Reinforcement Learning: Adaptivity and Computational Efficiency
    2023/02/21 by Heyang Zhao, Zhao, Heyang, Jiafan He +7 · 7 citations
    Computer Science · Decision Sciences · Engineering · #Advanced Bandit Algorithms Research #Age of Information Optimization #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Optimization and Control (math.OC) #Smart Grid Energy Management
  4. Nearly Minimax Optimal Reinforcement Learning for Linear Markov Decision Processes
    2022/12/12 by Jiafan He, Heyang Zhao, He, Jiafan +5 · 3 citations
    Computer Science · Mathematics · Medicine · #Advanced Causal Inference Techniques #FOS: Computer and information sciences #FOS: Mathematics #Influenza Virus Research Studies #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Optimization and Control (math.OC) #Reinforcement Learning in Robotics
  5. Nearly Optimal Algorithms for Linear Contextual Bandits with Adversarial Corruptions
    2022/05/13 by Jiafan He, Dongruo Zhou, He, Jiafan +5 · 2 citations
    Decision Sciences · Computer Science · #Advanced Bandit Algorithms Research #Optimization and Search Problems #Stochastic Gradient Optimization Techniques
  6. Towards Robust Model-Based Reinforcement Learning Against Adversarial Corruption
    2024/02/14 by Chenlu Ye, Ye, Chenlu, Jiafan He +5 · 3 citations
    Computer Science · #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  7. A Nearly Optimal and Low-Switching Algorithm for Reinforcement Learning with General Function Approximation
    2023/11/26 by Heyang Zhao, Jiafan He, Zhao, Heyang +3 · 2 citations
    Computer Science · Decision Sciences · #Reinforcement Learning in Robotics #Advanced Bandit Algorithms Research #Optimization and Search Problems
  8. Nemotron-Cascade 2: Post-Training LLMs with Cascade RL and Multi-Domain On-Policy Distillation
    2026/03/19 by Zhuolin Yang, Zihan Liu, Yang Chen +14 · 3 voices
    Computer Science · #cs.CL #cs.AI #cs.LG
  9. Optimal Online Generalized Linear Regression with Stochastic Noise and Its Application to Heteroscedastic Bandits
    2022/02/28 by Heyang Zhao, Zhao, Heyang, Dongruo Zhou +5 · 1 citation
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Distributed Sensor Networks and Detection Algorithms #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Optimization and Control (math.OC)
  10. Pessimistic Nonlinear Least-Squares Value Iteration for Offline Reinforcement Learning
    2023/10/02 by Qiwei Di, Heyang Zhao, Di, Qiwei +5 · 1 citation
    Decision Sciences · Mathematics · Computer Science · #Advanced Bandit Algorithms Research #COVID-19 epidemiological studies #Reinforcement Learning in Robotics
  11. Uniform-PAC Bounds for Reinforcement Learning with Linear Function\n Approximation
    2021/06/22 by Jiafan He, Dongruo Zhou, He, Jiafan +3 · 1 citation
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Optimization and Control (math.OC) #Reinforcement Learning in Robotics
  12. Provably Efficient Representation Selection in Low-rank Markov Decision Processes: From Online to Offline RL
    2021/06/22 by Weitong Zhang, Zhang, Weitong, Jiafan He +7 · 3 citations
    Computer Science · Decision Sciences · Engineering · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Optimization and Control (math.OC) #Reinforcement Learning in Robotics #Smart Grid Energy Management