vix.ing · top · new · best · stats · spec

Chittepu, Yaswanth

  1. Scaling Laws for Reward Model Overoptimization in Direct Alignment Algorithms
    2024/06/05 by Rafael Rafailov, Rafailov, Rafael, Yaswanth Chittepu +13 · 19 citations
    Engineering · Decision Sciences · Computer Science · #Advanced Control Systems Optimization #Simulation Techniques and Applications #Statistical and Computational Modeling
  2. Reinforcement Learning from Human Feedback with High-Confidence Safety Constraints
    2025/06/09 by Yaswanth Chittepu, Chittepu, Yaswanth, Blossom Metevier +9 · 2 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Applications (stat.AP) #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Topic Modeling