vix.ing · top · new · best · stats · spec

Théophile Sautory

  1. Reinforcement Learning from LLM Feedback to Counteract Goal Misgeneralization
    2024/01/14 by Houda Nait El Barj, Barj, Houda Nait El, Théophile Sautory +1 · 2 citations
    Computer Science · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Topic Modeling