vix.ing · top · new · best · stats · spec

Odalric-Ambrym Maillard

  1. Tightening Exploration in Upper Confidence Reinforcement Learning
    2020/04/20 by Hippolyte Bourel, Bourel, Hippolyte, Odalric-Ambrym Maillard +3 · 6 citations
    Computer Science · Decision Sciences · #Adaptive Dynamic Programming Control #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Systems and Control (eess.SY) #electronic engineering #information engineering
  2. Variance-Aware Regret Bounds for Undiscounted Reinforcement Learning in MDPs
    2018/03/05 by Mohammad Sadegh Talebi, Odalric-Ambrym Maillard, Talebi, Mohammad Sadegh +1 · 4 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Adversarial Robustness in Machine Learning #Age of Information Optimization #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Reinforcement Learning in Robotics #Systems and Control (eess.SY) #electronic engineering #information engineering
  3. Efficient tracking of a growing number of experts
    2017/08/31 by Jaouad Mourtada, Odalric-Ambrym Maillard, Mourtada, Jaouad +1 · 2 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Mobile Crowdsensing and Crowdsourcing
  4. Indexed Minimum Empirical Divergence for Unimodal Bandits
    2021/12/02 by Hassan Saber, Pierre Ménard, Saber, Hassan +3 · 2 citations
    Decision Sciences · Computer Science · #Advanced Bandit Algorithms Research #Stochastic Gradient Optimization Techniques #Machine Learning and Algorithms
  5. CRIMED: Lower and Upper Bounds on Regret for Bandits with Unbounded Stochastic Corruption
    2023/09/28 by Shubhada Agrawal, Timothée Mathieu, Agrawal, Shubhada +5 · 1 citation
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Auction Theory and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Optimization and Search Problems
  6. Provably Efficient Exploration in Reward Machines with Low Regret
    2024/12/26 by Hippolyte Bourel, Anders Jönsson, Bourel, Hippolyte +7 · 1 citation
    Engineering · Computer Science · #Advanced Control Systems Optimization #Fault Detection and Control Systems #Fuzzy Logic and Control Systems
  7. The regret lower bound for communicating Markov Decision Processes
    2025/01/22 by Victor Boone, Odalric-Ambrym Maillard, Boone, Victor +1 · 2 citations
    Computer Science · #Bayesian Modeling and Causal Inference #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  8. Bregman Deviations of Generic Exponential Families
    2022/01/18 by Sayak Ray Chowdhury, Chowdhury, Sayak Ray, Patrick Saux +5 · 1 citation
    Mathematics · Physics and Astronomy · #Advanced Statistical Methods and Models #Mathematical Inequalities and Applications #Statistical Mechanics and Entropy
  9. Kriging and Gaussian Process Interpolation for Georeferenced Data Augmentation
    2025/01/13 by Frédérick Fabre Ferber, Dominique Gay, Ferber, Frédérick Fabre +7 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Gaussian Processes and Bayesian Inference
  10. Information-Based Exploration via Random Features for Reinforcement Learning
    2026/07/20 by Waris Radji, Odalric-Ambrym Maillard
    #cs.LG