vix.ing · top · new · best · stats · spec

B. M. Wu

  1. LlamaRL: A Distributed Asynchronous Reinforcement Learning Framework for Efficient Large-scale LLM Training
    2025/05/29 by B. M. Wu, Wu, Bo, S Wang +25 · 17 citations
    Computer Science · Engineering · #Artificial Intelligence (cs.AI) #Elevator Systems and Control #FOS: Computer and information sciences #Fuzzy Logic and Control Systems #Machine Learning (cs.LG) #Speech and dialogue systems