vix.ing · top · new · best · stats · spec

Wan, Xiaopei

  1. Every Step Evolves: Scaling Reinforcement Learning for Trillion-Scale Thinking Model
    2025/10/21 by Ling Team, Shen, Anqi, Li, Baihui +101 · 2 voices · 7 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
  2. Ring-lite: Scalable Reasoning via C3PO-Stabilized Reinforcement Learning for LLMs
    2025/06/17 by Bin‐Jie Hu, Ling Team, Cai Chen +85 · 6 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Multi-Agent Systems and Negotiation #Natural Language Processing Techniques #Semantic Web and Ontologies