vix.ing · top · new · best · stats · spec

Bensal, Shelly

  1. Reflect, Retry, Reward: Self-Improving LLMs via Reinforcement Learning
    2025/05/30 by Bensal, Shelly, Jamil, Umar, Bryant, Christopher +5 · 12 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences