vix.ing · top · new · best · stats · spec

Ghavamzadeh, Mohammad

  1. DPOK: Reinforcement Learning for Fine-tuning Text-to-Image Diffusion Models
    2023/05/25 by Ying Fan, Olivia Watkins, Fan, Ying +17 · 102 citations
    Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG)
  2. Aligning Text-to-Image Models using Human Feedback
    2023/02/23 by Kimin Lee, Lee, Kimin, Hao Liu +15 · 70 citations
    Computer Science · #Advanced Image and Video Retrieval Techniques #Artificial Intelligence (cs.AI) #Augmented Reality Applications #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Handwritten Text Recognition Techniques #Machine Learning (cs.LG)
  3. Risk-Constrained Reinforcement Learning with Percentile Risk Criteria
    2015/12/05 by Yinlam Chow, Mohammad Ghavamzadeh, Chow, Yinlam +5 · 31 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Optimization and Control (math.OC) #Reinforcement Learning in Robotics #Risk and Portfolio Optimization
  4. A Lyapunov-based Approach to Safe Reinforcement Learning
    2018/05/20 by Yinlam Chow, Chow, Yinlam, Ofir Nachum +5 · 31 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Formal Methods in Verification #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
  5. Algorithms for CVaR Optimization in MDPs
    2014/06/12 by Yinlam Chow, Chow, Yinlam, Mohammad Ghavamzadeh +1 · 13 citations
    Decision Sciences · Engineering · #Artificial Intelligence (cs.AI) #Auction Theory and Applications #FOS: Computer and information sciences #FOS: Mathematics #Optimization and Control (math.OC) #Reservoir Engineering and Simulation Methods #Risk and Portfolio Optimization
  6. Policy Gradient for Coherent Risk Measures
    2015/02/13 by Aviv Tamar, Tamar, Aviv, Yinlam Chow +5 · 14 citations
    Decision Sciences · Computer Science · #Risk and Portfolio Optimization #Reinforcement Learning in Robotics #Advanced Bandit Algorithms Research
  7. Conservative Contextual Linear Bandits
    2016/11/19 by Abbas Kazerouni, Kazerouni, Abbas, Mohammad Ghavamzadeh +5 · 15 citations
    Decision Sciences · Engineering · Computer Science · #Advanced Bandit Algorithms Research #Smart Grid Energy Management #Machine Learning and Algorithms
  8. Lyapunov-based Safe Policy Optimization for Continuous Control
    2019/01/28 by Yinlam Chow, Chow, Yinlam, Ofir Nachum +7 · 11 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Robotic Path Planning Algorithms
  9. Robust Reinforcement Learning using Offline Data
    2022/08/10 by Kishan Panaganti, Zaiyan Xu, Panaganti, Kishan +5 · 14 citations
    Computer Science · Engineering · #Adversarial Robustness in Machine Learning #Reinforcement Learning in Robotics #Fault Detection and Control Systems
  10. Benchmarking Batch Deep Reinforcement Learning Algorithms
    2019/10/03 by Scott Fujimoto, Fujimoto, Scott, Edoardo Conti +5 · 17 citations
    Computer Science · Engineering · #Adaptive Dynamic Programming Control #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Smart Grid Energy Management
  11. More Robust Doubly Robust Off-policy Evaluation
    2018/02/10 by Mehrdad Farajtabar, Farajtabar, Mehrdad, Yinlam Chow +3 · 8 citations
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Reinforcement Learning in Robotics
  12. Upper-Confidence-Bound Algorithms for Active Learning in Multi-Armed Bandits
    2015/07/16 by Alexandra Carpentier, Alessandro Lazaric, Carpentier, Alexandra +9 · 6 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Data Stream Mining Techniques #FOS: Computer and information sciences #G.3 #Machine Learning (cs.LG) #Machine Learning and Algorithms
  13. Randomized Exploration in Generalized Linear Bandits
    2019/06/21 by Branislav Kveton, Manzil Zaheer, Kveton, Branislav +9 · 7 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Stochastic Gradient Optimization Techniques
  14. Mirror Descent Policy Optimization
    2020/05/20 by Tomar, Manan, Shani, Lior, Efroni, Yonathan +1 · 7 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  15. Safe Policy Improvement by Minimizing Robust Baseline Regret
    2016/07/13 by Marek Petrik, Petrik, Marek, Yinlam Chow +3 · 6 citations
    Decision Sciences · Engineering · #Advanced Control Systems Optimization #FOS: Computer and information sciences #Fault Detection and Control Systems #Machine Learning (stat.ML) #Probabilistic and Robust Engineering Design
  16. Robust Locally-Linear Controllable Embedding
    2017/10/15 by Ershad Banijamali, Rui Shu, Banijamali, Ershad +7 · 7 citations
    Computer Science · Engineering · Physics and Astronomy · #Control Systems and Identification #FOS: Computer and information sciences #Machine Learning (cs.LG) #Model Reduction and Neural Networks #Reinforcement Learning in Robotics
  17. Stochastic Bandits with Linear Constraints
    2020/06/17 by Aldo Pacchiano, Pacchiano, Aldo, Mohammad Ghavamzadeh +5 · 6 citations
    Computer Science · Decision Sciences · Engineering · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Optimization and Search Problems #Smart Grid Energy Management
  18. Efficient Risk-Averse Reinforcement Learning
    2022/05/10 by Ido Greenberg, Greenberg, Ido, Yinlam Chow +5 · 6 citations
    Computer Science · #Adversarial Robustness in Machine Learning #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
  19. A Review of Deep Learning for Video Captioning
    2023/04/22 by Moloud Abdar, Abdar, Moloud, Meenakshi Kollati +18 · 7 citations
    Computer Science · #Multimodal Machine Learning Applications #Human Pose and Action Recognition #Video Analysis and Summarization
  20. The Amazon Nova Family of Models: Technical Report and Model Card
    2025/03/17 by Amazon AGI, AGI, Amazon, Aayush Shah +867 · 21 citations
    Computer Science · Engineering · #3D Modeling in Geospatial Applications #Artificial Intelligence (cs.AI) #BIM and Construction Integration #Computers and Society (cs.CY) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Model-Driven Software Engineering Techniques
  21. Risk-Sensitive Generative Adversarial Imitation Learning
    2018/08/13 by Jonathan Lacotte, Mohammad Ghavamzadeh, Lacotte, Jonathan +5 · 5 citations
    Computer Science · Physics and Astronomy · #Anomaly Detection Techniques and Applications #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Human Pose and Action Recognition #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Model Reduction and Neural Networks
  22. Online Learning to Rank in Stochastic Click Models
    2017/03/07 by Zoghi, Masrour, Tunys, Tomas, Ghavamzadeh, Mohammad +3 · 3 citations
    #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  23. Perturbed-History Exploration in Stochastic Linear Bandits
    2019/03/21 by Kveton, Branislav, Szepesvari, Csaba, Ghavamzadeh, Mohammad +1 · 3 citations
    #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  24. Prediction, Consistency, Curvature: Representation Learning for Locally-Linear Control
    2019/09/04 by Levine, Nir, Chow, Yinlam, Shu, Rui +3 · 3 citations
    #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  25. Hierarchical Bayesian Bandits
    2021/11/12 by Hong, Joey, Kveton, Branislav, Zaheer, Manzil +1 · 3 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  26. Path Consistency Learning in Tsallis Entropy Regularized MDPs
    2018/02/10 by Ofir Nachum, Nachum, Ofir, Yinlam Chow +3 · 2 citations
    Computer Science · Decision Sciences · #Adaptive Dynamic Programming Control #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Probabilistic and Robust Engineering Design #Reinforcement Learning in Robotics
  27. Predictive Coding for Locally-Linear Control
    2020/03/02 by Shu, Rui, Nguyen, Tung, Chow, Yinlam +5 · 2 citations
    #FOS: Computer and information sciences #FOS: Electrical engineering #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Systems and Control (eess.SY) #electronic engineering #information engineering
  28. Does Thinking More always Help? Mirage of Test-Time Scaling in Reasoning Models
    2025/06/04 by Soumya Suvra Ghosal, Ghosal, Soumya Suvra, Avinash Reddy +14 · 13 citations
    Computer Science · Social Sciences · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Ethics and Social Impacts of AI #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Mobile Crowdsensing and Crowdsourcing
  29. A Dantzig Selector Approach to Temporal Difference Learning
    2012/06/27 by Matthieu Geist, Bruno Scherrer, Geist, Matthieu +5 · 2 citations
    Biochemistry, Genetics and Molecular Biology · Computer Science · #Cancer-related molecular mechanisms research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Music and Audio Processing #Speech and Audio Processing
  30. Tight Regret Bounds for Model-Based Reinforcement Learning with Greedy\n Policies
    2019/05/27 by Yonathan Efroni, Nadav Merlis, Efroni, Yonathan +5 · 3 citations
    Computer Science · Decision Sciences · Engineering · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #Explainable Artificial Intelligence (XAI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Smart Grid Energy Management
  31. Bridging Distributionally Robust Learning and Offline RL: An Approach to Mitigate Distribution Shift and Partial Data Coverage
    2023/10/27 by Kishan Panaganti, Panaganti, Kishan, Zaiyan Xu +5 · 2 citations
    Computer Science · Decision Sciences · #Adaptive Dynamic Programming Control #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
  32. Factual and Personalized Recommendations using Language Models and Reinforcement Learning
    2023/10/09 by Jeong, Jihwan, Chow, Yinlam, Tennenholtz, Guy +4 · 2 citations
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences
  33. Finite-Sample Analysis of Proximal Gradient TD Algorithms
    2020/06/06 by Bo Liu, Ji Liu, Liu, Bo +7 · 4 citations
    Computer Science · Engineering · #Adaptive Dynamic Programming Control #FOS: Computer and information sciences #Iterative Learning Control Systems #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
  34. Active Model Estimation in Markov Decision Processes
    2020/03/06 by Tarbouriech, Jean, Shekhar, Shubhanshu, Pirotta, Matteo +2 · 1 citation
    #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  35. Policy-Aware Model Learning for Policy Gradient Methods
    2020/02/28 by Abachi, Romina, Ghavamzadeh, Mohammad, Farahmand, Amir-massoud · 1 citation
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences
  36. Variational Model-based Policy Optimization
    2020/06/09 by Chow, Yinlam, Cui, Brandon, Ryu, MoonKyung +1 · 1 citation
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  37. Non-Stationary Latent Bandits
    2020/12/01 by Joey Hong, Hong, Joey, Branislav Kveton +11 · 1 citation
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #Data Stream Mining Techniques #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Algorithms
  38. Adaptive Sampling for Minimax Fair Classification
    2021/03/01 by Shubhanshu Shekhar, Shekhar, Shubhanshu, Fields, Greg +4 · 1 citation
    Computer Science · Social Sciences · #Ethics and Social Impacts of AI #FOS: Computer and information sciences #Imbalanced Data Classification Techniques #Machine Learning (cs.LG) #Privacy-Preserving Technologies in Data
  39. Robust Policy Optimization with Baseline Guarantees
    2015/06/15 by Yinlam Chow, Marek Petrik, Chow, Yinlam +3 · 1 citation
    Computer Science · Decision Sciences · #FOS: Mathematics #Machine Learning and Algorithms #Optimization and Control (math.OC) #Reinforcement Learning in Robotics #Risk and Portfolio Optimization
  40. Fixed-Budget Best-Arm Identification in Structured Bandits
    2021/06/09 by Mohammad Javad Azizi, Branislav Kveton, Azizi, Mohammad Javad +3 · 1 citation
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Auction Theory and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Algorithms
  41. Meta-Learning for Simple Regret Minimization
    2022/02/25 by Mohammadjavad Azizi, Azizi, Mohammadjavad, Branislav Kveton +5 · 1 citation
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Machine Learning and Data Classification
  42. Multi-Task Off-Policy Learning from Bandit Feedback
    2022/12/09 by Hong, Joey, Kveton, Branislav, Katariya, Sumeet +2 · 1 citation
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG)
  43. Q-learning for Quantile MDPs: A Decomposition, Performance, and Convergence Analysis
    2024/10/31 by Hau, Jia Lin, Delage, Erick, Derman, Esther +2 · 2 citations
    #FOS: Computer and information sciences #Machine Learning (cs.LG)
  44. Offline Reinforcement Learning for Mixture-of-Expert Dialogue Management
    2023/02/21 by Dhawal Gupta, Gupta, Dhawal, Yinlam Chow +6 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Speech and dialogue systems #Topic Modeling
  45. Contextual Bandits with Stage-wise Constraints
    2024/01/15 by Pacchiano, Aldo, Ghavamzadeh, Mohammad, Bartlett, Peter · 1 citation
    #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  46. Preference Elicitation with Soft Attributes in Interactive Recommendation
    2023/10/22 by Biyik, Erdem, Yao, Fan, Chow, Yinlam +4 · 1 citation
    #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Information Retrieval (cs.IR)
  47. Conservative Contextual Bandits: Beyond Linear Representations
    2024/12/09 by Rohan Deb, Deb, Rohan, Mohammad Ghavamzadeh +3 · 1 citation
    Decision Sciences · Social Sciences · #Decision-Making and Behavioral Economics #Misinformation and Its Impacts #Experimental Behavioral Economics Studies