Ghavamzadeh, Mohammad
- DPOK: Reinforcement Learning for Fine-tuning Text-to-Image Diffusion Models
2023/05/25 by Ying Fan, Olivia Watkins, Fan, Ying +17 · 102 citations
Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG)
- Aligning Text-to-Image Models using Human Feedback
2023/02/23 by Kimin Lee, Lee, Kimin, Hao Liu +15 · 70 citations
Computer Science · #Advanced Image and Video Retrieval Techniques #Artificial Intelligence (cs.AI) #Augmented Reality Applications #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Handwritten Text Recognition Techniques #Machine Learning (cs.LG)
- Risk-Constrained Reinforcement Learning with Percentile Risk Criteria
2015/12/05 by Yinlam Chow, Mohammad Ghavamzadeh, Chow, Yinlam +5 · 31 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Optimization and Control (math.OC) #Reinforcement Learning in Robotics #Risk and Portfolio Optimization
- A Lyapunov-based Approach to Safe Reinforcement Learning
2018/05/20 by Yinlam Chow, Chow, Yinlam, Ofir Nachum +5 · 31 citations
Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Formal Methods in Verification #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
- Algorithms for CVaR Optimization in MDPs
2014/06/12 by Yinlam Chow, Chow, Yinlam, Mohammad Ghavamzadeh +1 · 13 citations
Decision Sciences · Engineering · #Artificial Intelligence (cs.AI) #Auction Theory and Applications #FOS: Computer and information sciences #FOS: Mathematics #Optimization and Control (math.OC) #Reservoir Engineering and Simulation Methods #Risk and Portfolio Optimization
- Policy Gradient for Coherent Risk Measures
2015/02/13 by Aviv Tamar, Tamar, Aviv, Yinlam Chow +5 · 14 citations
Decision Sciences · Computer Science · #Risk and Portfolio Optimization #Reinforcement Learning in Robotics #Advanced Bandit Algorithms Research
- Conservative Contextual Linear Bandits
2016/11/19 by Abbas Kazerouni, Kazerouni, Abbas, Mohammad Ghavamzadeh +5 · 15 citations
Decision Sciences · Engineering · Computer Science · #Advanced Bandit Algorithms Research #Smart Grid Energy Management #Machine Learning and Algorithms
- Lyapunov-based Safe Policy Optimization for Continuous Control
2019/01/28 by Yinlam Chow, Chow, Yinlam, Ofir Nachum +7 · 11 citations
Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Robotic Path Planning Algorithms
- Robust Reinforcement Learning using Offline Data
2022/08/10 by Kishan Panaganti, Zaiyan Xu, Panaganti, Kishan +5 · 14 citations
Computer Science · Engineering · #Adversarial Robustness in Machine Learning #Reinforcement Learning in Robotics #Fault Detection and Control Systems
- Benchmarking Batch Deep Reinforcement Learning Algorithms
2019/10/03 by Scott Fujimoto, Fujimoto, Scott, Edoardo Conti +5 · 17 citations
Computer Science · Engineering · #Adaptive Dynamic Programming Control #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Smart Grid Energy Management
- More Robust Doubly Robust Off-policy Evaluation
2018/02/10 by Mehrdad Farajtabar, Farajtabar, Mehrdad, Yinlam Chow +3 · 8 citations
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Reinforcement Learning in Robotics
- Upper-Confidence-Bound Algorithms for Active Learning in Multi-Armed Bandits
2015/07/16 by Alexandra Carpentier, Alessandro Lazaric, Carpentier, Alexandra +9 · 6 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Data Stream Mining Techniques #FOS: Computer and information sciences #G.3 #Machine Learning (cs.LG) #Machine Learning and Algorithms
- Randomized Exploration in Generalized Linear Bandits
2019/06/21 by Branislav Kveton, Manzil Zaheer, Kveton, Branislav +9 · 7 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Stochastic Gradient Optimization Techniques
- Mirror Descent Policy Optimization
2020/05/20 by Tomar, Manan, Shani, Lior, Efroni, Yonathan +1 · 7 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Safe Policy Improvement by Minimizing Robust Baseline Regret
2016/07/13 by Marek Petrik, Petrik, Marek, Yinlam Chow +3 · 6 citations
Decision Sciences · Engineering · #Advanced Control Systems Optimization #FOS: Computer and information sciences #Fault Detection and Control Systems #Machine Learning (stat.ML) #Probabilistic and Robust Engineering Design
- Robust Locally-Linear Controllable Embedding
2017/10/15 by Ershad Banijamali, Rui Shu, Banijamali, Ershad +7 · 7 citations
Computer Science · Engineering · Physics and Astronomy · #Control Systems and Identification #FOS: Computer and information sciences #Machine Learning (cs.LG) #Model Reduction and Neural Networks #Reinforcement Learning in Robotics
- Stochastic Bandits with Linear Constraints
2020/06/17 by Aldo Pacchiano, Pacchiano, Aldo, Mohammad Ghavamzadeh +5 · 6 citations
Computer Science · Decision Sciences · Engineering · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Optimization and Search Problems #Smart Grid Energy Management
- Efficient Risk-Averse Reinforcement Learning
2022/05/10 by Ido Greenberg, Greenberg, Ido, Yinlam Chow +5 · 6 citations
Computer Science · #Adversarial Robustness in Machine Learning #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
- A Review of Deep Learning for Video Captioning
2023/04/22 by Moloud Abdar, Abdar, Moloud, Meenakshi Kollati +18 · 7 citations
Computer Science · #Multimodal Machine Learning Applications #Human Pose and Action Recognition #Video Analysis and Summarization
- The Amazon Nova Family of Models: Technical Report and Model Card
2025/03/17 by Amazon AGI, AGI, Amazon, Aayush Shah +867 · 21 citations
Computer Science · Engineering · #3D Modeling in Geospatial Applications #Artificial Intelligence (cs.AI) #BIM and Construction Integration #Computers and Society (cs.CY) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Model-Driven Software Engineering Techniques
- Risk-Sensitive Generative Adversarial Imitation Learning
2018/08/13 by Jonathan Lacotte, Mohammad Ghavamzadeh, Lacotte, Jonathan +5 · 5 citations
Computer Science · Physics and Astronomy · #Anomaly Detection Techniques and Applications #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Human Pose and Action Recognition #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Model Reduction and Neural Networks
- Online Learning to Rank in Stochastic Click Models
2017/03/07 by Zoghi, Masrour, Tunys, Tomas, Ghavamzadeh, Mohammad +3 · 3 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Perturbed-History Exploration in Stochastic Linear Bandits
2019/03/21 by Kveton, Branislav, Szepesvari, Csaba, Ghavamzadeh, Mohammad +1 · 3 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Prediction, Consistency, Curvature: Representation Learning for Locally-Linear Control
2019/09/04 by Levine, Nir, Chow, Yinlam, Shu, Rui +3 · 3 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Hierarchical Bayesian Bandits
2021/11/12 by Hong, Joey, Kveton, Branislav, Zaheer, Manzil +1 · 3 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Path Consistency Learning in Tsallis Entropy Regularized MDPs
2018/02/10 by Ofir Nachum, Nachum, Ofir, Yinlam Chow +3 · 2 citations
Computer Science · Decision Sciences · #Adaptive Dynamic Programming Control #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Probabilistic and Robust Engineering Design #Reinforcement Learning in Robotics
- Predictive Coding for Locally-Linear Control
2020/03/02 by Shu, Rui, Nguyen, Tung, Chow, Yinlam +5 · 2 citations
#FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Systems and Control (eess.SY) #electronic engineering #information engineering
- Does Thinking More always Help? Mirage of Test-Time Scaling in Reasoning Models
2025/06/04 by Soumya Suvra Ghosal, Ghosal, Soumya Suvra, Avinash Reddy +14 · 13 citations
Computer Science · Social Sciences · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Mobile Crowdsensing and Crowdsourcing
- A Dantzig Selector Approach to Temporal Difference Learning
2012/06/27 by Matthieu Geist, Bruno Scherrer, Geist, Matthieu +5 · 2 citations
Biochemistry, Genetics and Molecular Biology · Computer Science · #Cancer-related molecular mechanisms research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music and Audio Processing #Speech and Audio Processing
- Tight Regret Bounds for Model-Based Reinforcement Learning with Greedy\n Policies
2019/05/27 by Yonathan Efroni, Nadav Merlis, Efroni, Yonathan +5 · 3 citations
Computer Science · Decision Sciences · Engineering · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Smart Grid Energy Management
- Bridging Distributionally Robust Learning and Offline RL: An Approach to Mitigate Distribution Shift and Partial Data Coverage
2023/10/27 by Kishan Panaganti, Panaganti, Kishan, Zaiyan Xu +5 · 2 citations
Computer Science · Decision Sciences · #Adaptive Dynamic Programming Control #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
- Factual and Personalized Recommendations using Language Models and Reinforcement Learning
2023/10/09 by Jeong, Jihwan, Chow, Yinlam, Tennenholtz, Guy +4 · 2 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences
- Finite-Sample Analysis of Proximal Gradient TD Algorithms
2020/06/06 by Bo Liu, Ji Liu, Liu, Bo +7 · 4 citations
Computer Science · Engineering · #Adaptive Dynamic Programming Control #FOS: Computer and information sciences #Iterative Learning Control Systems #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
- Active Model Estimation in Markov Decision Processes
2020/03/06 by Tarbouriech, Jean, Shekhar, Shubhanshu, Pirotta, Matteo +2 · 1 citation
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Policy-Aware Model Learning for Policy Gradient Methods
2020/02/28 by Abachi, Romina, Ghavamzadeh, Mohammad, Farahmand, Amir-massoud · 1 citation
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences
- Variational Model-based Policy Optimization
2020/06/09 by Chow, Yinlam, Cui, Brandon, Ryu, MoonKyung +1 · 1 citation
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Non-Stationary Latent Bandits
2020/12/01 by Joey Hong, Hong, Joey, Branislav Kveton +11 · 1 citation
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #Data Stream Mining Techniques #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Algorithms
- Adaptive Sampling for Minimax Fair Classification
2021/03/01 by Shubhanshu Shekhar, Shekhar, Shubhanshu, Fields, Greg +4 · 1 citation
Computer Science · Social Sciences · #Ethics and Social Impacts of AI #FOS: Computer and information sciences #Imbalanced Data Classification Techniques #Machine Learning (cs.LG) #Privacy-Preserving Technologies in Data
- Robust Policy Optimization with Baseline Guarantees
2015/06/15 by Yinlam Chow, Marek Petrik, Chow, Yinlam +3 · 1 citation
Computer Science · Decision Sciences · #FOS: Mathematics #Machine Learning and Algorithms #Optimization and Control (math.OC) #Reinforcement Learning in Robotics #Risk and Portfolio Optimization
- Fixed-Budget Best-Arm Identification in Structured Bandits
2021/06/09 by Mohammad Javad Azizi, Branislav Kveton, Azizi, Mohammad Javad +3 · 1 citation
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Auction Theory and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Algorithms
- Meta-Learning for Simple Regret Minimization
2022/02/25 by Mohammadjavad Azizi, Azizi, Mohammadjavad, Branislav Kveton +5 · 1 citation
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Machine Learning and Data Classification
- Multi-Task Off-Policy Learning from Bandit Feedback
2022/12/09 by Hong, Joey, Kveton, Branislav, Katariya, Sumeet +2 · 1 citation
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Q-learning for Quantile MDPs: A Decomposition, Performance, and Convergence Analysis
2024/10/31 by Hau, Jia Lin, Delage, Erick, Derman, Esther +2 · 2 citations
#FOS: Computer and information sciences #Machine Learning (cs.LG)
- Offline Reinforcement Learning for Mixture-of-Expert Dialogue Management
2023/02/21 by Dhawal Gupta, Gupta, Dhawal, Yinlam Chow +6 · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Speech and dialogue systems #Topic Modeling
- Contextual Bandits with Stage-wise Constraints
2024/01/15 by Pacchiano, Aldo, Ghavamzadeh, Mohammad, Bartlett, Peter · 1 citation
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Preference Elicitation with Soft Attributes in Interactive Recommendation
2023/10/22 by Biyik, Erdem, Yao, Fan, Chow, Yinlam +4 · 1 citation
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Information Retrieval (cs.IR)
- Conservative Contextual Bandits: Beyond Linear Representations
2024/12/09 by Rohan Deb, Deb, Rohan, Mohammad Ghavamzadeh +3 · 1 citation
Decision Sciences · Social Sciences · #Decision-Making and Behavioral Economics #Misinformation and Its Impacts #Experimental Behavioral Economics Studies