Rohan Surana
- In-context Ranking Preference Optimization
2025/04/21 by Junda Wu, Rohan Surana, Wu, Junda +15 · 2 voices · 2 citations
Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #cs.LG
- Can We Break LLMs Out of Self-Loops? Fine-Grained Reasoning Control with Activation Steering
2026/07/20 by Sheldon Yu, Tong Yu, Xunyi Jiang +6
#cs.AI
- RRPO: Reference-Relative Policy Optimization with Stratified Conditional Rollouts
2026/07/20 by Yuxin Xiong, Xunyi Jiang, Rohan Surana +8
#cs.LG #cs.AI