vix.ing
·
top
·
new
·
best
·
stats
·
spec
Bensal, Shelly
Reflect, Retry, Reward: Self-Improving LLMs via Reinforcement Learning
2025/05/30 by
Bensal, Shelly
,
Jamil, Umar
,
Bryant, Christopher
+5 · 12 citations
#Computation and Language (cs.CL)
#FOS: Computer and information sciences