vix.ing · top · new · best · stats · spec

Lan, Qingfeng

  1. Maxmin Q-learning: Controlling the Estimation Bias of Q-learning
    2020/02/16 by Qingfeng Lan, Yangchen Pan, Lan, Qingfeng +5 · 15 citations
    Computer Science · #Reinforcement Learning in Robotics #Domain Adaptation and Few-Shot Learning #Adversarial Robustness in Machine Learning
  2. Variational Quantum Soft Actor-Critic
    2021/12/20 by Qingfeng Lan, Lan, Qingfeng · 3 citations
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #FOS: Physical sciences #Machine Learning (cs.LG) #Neural Networks and Reservoir Computing #Quantum Computing Algorithms and Architecture #Quantum Information and Cryptography #Quantum Physics (quant-ph)
  3. Weight Clipping for Deep Continual and Reinforcement Learning
    2024/07/01 by Elsayed, Mohamed, Lan, Qingfeng, Lyle, Clare +1 · 6 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  4. Provable and Practical: Efficient Exploration in Reinforcement Learning via Langevin Monte Carlo
    2023/05/29 by Ishfaq, Haque, Lan, Qingfeng, Xu, Pan +4 · 3 citations
    #FOS: Computer and information sciences #Machine Learning (cs.LG)
  5. Memory-efficient Reinforcement Learning with Value-based Knowledge Consolidation
    2022/05/22 by Lan, Qingfeng, Pan, Yangchen, Luo, Jun +1 · 2 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  6. Model-free Policy Learning with Reward Gradients
    2021/03/09 by Lan, Qingfeng, Tosatto, Samuele, Farrahi, Homayoon +1 · 1 citation
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  7. Learning to Optimize for Reinforcement Learning
    2023/02/03 by Lan, Qingfeng, Mahmood, A. Rupam, Yan, Shuicheng +1 · 1 citation
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  8. More Efficient Randomized Exploration for Reinforcement Learning via Approximate Sampling
    2024/06/18 by Haque Ishfaq, Ishfaq, Haque, Yixin Tan +13 · 2 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #Distributed Sensor Networks and Detection Algorithms #Energy Efficient Wireless Sensor Networks #FOS: Computer and information sciences #Machine Learning (cs.LG)