vix.ing · top · new · best · stats · spec

Zhiyu Mei

  1. AReaL: A Large-Scale Asynchronous Reinforcement Learning System for Language Reasoning
    2025/05/30 by Fu Wei, Fu, Wei, Jiaxuan Gao +21 · 67 citations
    Computer Science · #Natural Language Processing Techniques #Fuzzy Logic and Control Systems #Speech and dialogue systems
  2. On Designing Effective RL Reward at Training Time for LLM Reasoning
    2024/10/19 by Jiaxuan Gao, Gao, Jiaxuan, Shusheng Xu +15 · 24 citations
    Computer Science · #Software Engineering Research
  3. Beyond Ten Turns: Unlocking Long-Horizon Agentic Search with Large-Scale Asynchronous RL
    2025/08/11 by Jiaxuan Gao, Gao, Jiaxuan, Wei Fu +13 · 52 citations
    Computer Science · #Robotic Path Planning Algorithms #Optimization and Search Problems #Reinforcement Learning in Robotics
  4. ReaL: Efficient RLHF Training of Large Language Models with Parameter Reallocation
    2024/06/20 by Zhiyu Mei, Mei, Zhiyu, Wei Fu +9 · 10 citations
    Computer Science · #Natural Language Processing Techniques #Topic Modeling #Speech Recognition and Synthesis
  5. SRL: Scaling Distributed Reinforcement Learning to Over Ten Thousand Cores
    2023/06/29 by Zhiyu Mei, Mei, Zhiyu, Wei Fu +8 · 2 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Distributed #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Parallel #Reinforcement Learning in Robotics #Software Engineering Research #and Cluster Computing (cs.DC)