vix.ing · top · new · best · stats · spec

Xue, Wanqi

  1. PrefRec: Recommender Systems with Human Preferences for Reinforcing Long-term User Engagement
    2022/12/06 by Wanqi Xue, Qingpeng Cai, Xue, Wanqi +13 · 5 citations
    Computer Science · Decision Sciences · #Recommender Systems and Techniques #Advanced Bandit Algorithms Research #Reinforcement Learning in Robotics
  2. Solving Large-Scale Extensive-Form Network Security Games via Neural Fictitious Self-Play
    2021/06/02 by Xue, Wanqi, Zhang, Youzhi, Li, Shuxin +3 · 3 citations
    #Artificial Intelligence (cs.AI) #Computer Science and Game Theory (cs.GT) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multiagent Systems (cs.MA)
  3. NSGZero: Efficiently Learning Non-Exploitable Policy in Large-Scale Network Security Games with Neural Monte Carlo Tree Search
    2022/01/17 by Xue, Wanqi, An, Bo, Yeo, Chai Kiat · 2 citations
    #Artificial Intelligence (cs.AI) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multiagent Systems (cs.MA)
  4. Scaling New Frontiers: Insights into Large Recommendation Models
    2024/12/01 by Guo, Wei, Wang, Hao, Zhang, Luankang +16 · 5 citations
    #FOS: Computer and information sciences #Information Retrieval (cs.IR)
  5. CFR-MIX: Solving Imperfect Information Extensive-Form Games with Combinatorial Action Space
    2021/05/18 by Li, Shuxin, Zhang, Youzhi, Wang, Xinrun +2 · 1 citation
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences
  6. Mis-spoke or mis-lead: Achieving Robustness in Multi-Agent Communicative Reinforcement Learning
    2021/08/09 by Xue, Wanqi, Qiu, Wei, An, Bo +3 · 1 citation
    #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multiagent Systems (cs.MA)
  7. ResAct: Reinforcing Long-term Engagement in Sequential Recommendation with Residual Actor
    2022/06/01 by Xue, Wanqi, Cai, Qingpeng, Zhan, Ruohan +4 · 1 citation
    #FOS: Computer and information sciences #Information Retrieval (cs.IR) #Machine Learning (cs.LG)