vix.ing · top · new · best · stats · spec

Makarova, Anastasiia

  1. RRM: Robust Reward Model Training Mitigates Reward Hacking
    2024/09/20 by Tianqi Liu, Liu, Tianqi, Wei Xiong +33 · 30 citations
    Psychology · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Human-Automation Interaction and Safety
  2. Model-based Causal Bayesian Optimization
    2022/11/18 by Scott Sussex, Anastasiia Makarova, Sussex, Scott +3 · 3 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Gaussian Processes and Bayesian Inference #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms