Will Schwarzer
- Reinforcement Learning from Human Feedback with High-Confidence Safety Constraints
2025/06/09 by Yaswanth Chittepu, Chittepu, Yaswanth, Blossom Metevier +9 · 2 citations
Computer Science · #Adversarial Robustness in Machine Learning #Applications (stat.AP) #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Topic Modeling