vix.ing · top · new · best · stats · spec

Qihong Lin

  1. ACRL: Adaptive Control of Training-Inference Discrepancy for Stable Reinforcement Learning
    2026/07/27 by Wenwu Fan, Qihong Lin, Zhijie Xia +4
    #cs.LG #cs.AI #cs.CL