Mohammad Sadegh Talebi
- Tightening Exploration in Upper Confidence Reinforcement Learning
2020/04/20 by Hippolyte Bourel, Bourel, Hippolyte, Odalric-Ambrym Maillard +3 · 6 citations
Computer Science · Decision Sciences · #Adaptive Dynamic Programming Control #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Systems and Control (eess.SY) #electronic engineering #information engineering
- Variance-Aware Regret Bounds for Undiscounted Reinforcement Learning in MDPs
2018/03/05 by Mohammad Sadegh Talebi, Odalric-Ambrym Maillard, Talebi, Mohammad Sadegh +1 · 4 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Adversarial Robustness in Machine Learning #Age of Information Optimization #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Reinforcement Learning in Robotics #Systems and Control (eess.SY) #electronic engineering #information engineering
- Spectral geometry on manifolds with fibred boundary metrics II: heat kernel asymptotics
2021/01/21 by Mohammad Sadegh Talebi, Boris Vertman, Talebi, Mohammad +1 · 1 citation
Mathematics · #58J52 #Analysis of PDEs (math.AP) #Differential Geometry (math.DG) #FOS: Mathematics #Geometric Analysis and Curvature Flows #Geometry and complex manifolds #Mathematical Dynamics and Fractals
- Provably Efficient Exploration in Reward Machines with Low Regret
2024/12/26 by Hippolyte Bourel, Anders Jönsson, Bourel, Hippolyte +7 · 1 citation
Engineering · Computer Science · #Advanced Control Systems Optimization #Fault Detection and Control Systems #Fuzzy Logic and Control Systems