Masatoshi Uehara
- Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control
2024/02/23 by Masatoshi Uehara, Uehara, Masatoshi, Yulai Zhao +15 · 20 citations
Computer Science · Mathematics · Physics and Astronomy · #Advanced Mathematical Modeling in Engineering #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Model Reduction and Neural Networks #Numerical methods in inverse problems
- A Review of Off-Policy Evaluation in Reinforcement Learning
2022/12/13 by Masatoshi Uehara, Uehara, Masatoshi, Chengchun Shi +3 · 1 voice · 9 citations
Computer Science · #Reinforcement Learning in Robotics #cs.LG #math.ST #stat.ME #stat.ML
- Understanding Reinforcement Learning-Based Fine-Tuning of Diffusion Models: A Tutorial and Review
2024/07/18 by Masatoshi Uehara, Uehara, Masatoshi, Yulai Zhao +5 · 19 citations
Engineering · #Artificial Intelligence (cs.AI) #FOS: Biological sciences #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Quantitative Methods (q-bio.QM) #Traffic control and management
- Double Reinforcement Learning for Efficient Off-Policy Evaluation in Markov Decision Processes
2019/08/22 by Nathan Kallus, Masatoshi Uehara, Kallus, Nathan +1 · 11 citations
Mathematics · Computer Science · Biochemistry, Genetics and Molecular Biology · #Advanced Causal Inference Techniques #Reinforcement Learning in Robotics #Gene Regulatory Network Analysis
- Inference-Time Alignment in Diffusion Models with Reward-Guided Generation: Tutorial and Review
2025/01/16 by Masatoshi Uehara, Uehara, Masatoshi, Yulai Zhao +11 · 1 voice · 16 citations
Physics and Astronomy · #Model Reduction and Neural Networks #cs.AI #cs.LG #q-bio.QM #stat.ML
- Minimax Weight and Q-Function Learning for Off-Policy Evaluation
2019/10/28 by Masatoshi Uehara, Jiawei Huang, Uehara, Masatoshi +3 · 6 citations
Computer Science · #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Formal Methods in Verification #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
- Fine-Tuning Discrete Diffusion Models via Reward Optimization with Applications to DNA and Protein Design
2024/10/17 by Chenyu Wang, Masatoshi Uehara, Wang, Chenyu +17 · 15 citations
Biochemistry, Genetics and Molecular Biology · #DNA and Nucleic Acid Chemistry
- Provable Offline Preference-Based Reinforcement Learning
2023/05/24 by Wenhao Zhan, Masatoshi Uehara, Zhan, Wenhao +7 · 6 citations
Biochemistry, Genetics and Molecular Biology · Computer Science · #Receptor Mechanisms and Signaling #Reinforcement Learning in Robotics #Formal Methods in Verification
- Pessimistic Model-based Offline Reinforcement Learning under Partial Coverage
2021/07/13 by Masatoshi Uehara, Uehara, Masatoshi, W. Sun +1 · 4 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Reinforcement Learning in Robotics
- Feedback Efficient Online Fine-Tuning of Diffusion Models
2024/02/26 by Masatoshi Uehara, Yulai Zhao, Uehara, Masatoshi +15 · 6 citations
Mathematics · Computer Science · Physics and Astronomy · #Numerical methods for differential equations #Matrix Theory and Algorithms #Model Reduction and Neural Networks
- Localized Debiased Machine Learning: Efficient Inference on Quantile Treatment Effects and Beyond
2019/12/30 by Nathan Kallus, Kallus, Nathan, Xiaojie Mao +3 · 3 citations
Mathematics · #Advanced Causal Inference Techniques #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Methodology (stat.ME) #Statistical Methods and Bayesian Inference #Statistical Methods and Inference
- Dynamic Search for Inference-Time Alignment in Diffusion Models
2025/03/03 by Xiner Li, Li, Xiner, Masatoshi Uehara +13 · 11 citations
Physics and Astronomy · #Model Reduction and Neural Networks
- Provable Reward-Agnostic Preference-Based Reinforcement Learning
2023/05/29 by Wenhao Zhan, Zhan, Wenhao, Masatoshi Uehara +5 · 4 citations
Computer Science · #Advanced Graph Neural Networks #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Statistics Theory (math.ST)
- Representation Learning for Online and Offline RL in Low-rank MDPs
2021/10/09 by Masatoshi Uehara, Uehara, Masatoshi, Xuezhou Zhang +3 · 3 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Reinforcement Learning in Robotics
- Provably Efficient Reinforcement Learning in Partially Observable Dynamical Systems
2022/06/24 by Masatoshi Uehara, Ayush Sekhari, Uehara, Masatoshi +7 · 3 citations
Computer Science · #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Methodology (stat.ME) #Reinforcement Learning in Robotics #Statistics Theory (math.ST)
- Finite Sample Analysis of Minimax Offline Reinforcement Learning: Completeness, Fast Rates and First-Order Efficiency
2021/02/05 by Masatoshi Uehara, Masaaki Imaizumi, Uehara, Masatoshi +9 · 6 citations
Computer Science · #FOS: Computer and information sciences #FOS: Mathematics #Formal Methods in Verification #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Software Reliability and Analysis Research #Statistics Theory (math.ST)
- Future-Dependent Value-Based Off-Policy Evaluation in POMDPs
2022/07/26 by Masatoshi Uehara, Uehara, Masatoshi, Haruka Kiyohara +13 · 2 citations
Engineering · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Smart Grid Energy Management
- Off-Policy Evaluation and Learning for External Validity under a\n Covariate Shift
2020/02/26 by Masahiro Kato, Kato, Masahiro, Masatoshi Uehara +3 · 7 citations
Mathematics · Engineering · Economics, Econometrics and Finance · #Advanced Causal Inference Techniques #Nuclear reactor physics and engineering #Economic Policies and Impacts
- Reward-Guided Iterative Refinement in Diffusion Models at Test-Time with Applications to Protein and DNA Design
2025/02/20 by Masatoshi Uehara, Uehara, Masatoshi, Xingyu Su +13 · 1 voice · 3 citations
Biochemistry, Genetics and Molecular Biology · Computer Science · Physics and Astronomy · #Gene Regulatory Network Analysis #Generative Adversarial Networks and Image Synthesis #Model Reduction and Neural Networks
- Mitigating Covariate Shift in Imitation Learning via Offline Data Without Great Coverage
2021/06/06 by Jonathan Chang, Masatoshi Uehara, Chang, Jonathan D. +7 · 3 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Reinforcement Learning in Robotics
- Causal Inference Under Unmeasured Confounding With Negative Controls: A Minimax Learning Approach
2021/03/25 by Nathan Kallus, Xiaojie Mao, Kallus, Nathan +3 · 1 citation
Computer Science · Mathematics · #Advanced Causal Inference Techniques #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Methodology (stat.ME) #Statistical Methods and Inference
- Offline Minimax Soft-Q-learning Under Realizability and Partial Coverage
2023/02/05 by Masatoshi Uehara, Uehara, Masatoshi, Nathan Kallus +5 · 2 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
- Minimax Instrumental Variable Regression and L2 Convergence Guarantees without Identification or Closedness
2023/02/10 by Andrew F. Bennett, Nathan Kallus, Bennett, Andrew +9 · 1 citation
Computer Science · Mathematics · #Distributed Sensor Networks and Detection Algorithms #Econometrics (econ.EM) #FOS: Computer and information sciences #FOS: Economics and business #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Methodology (stat.ME) #Statistical Methods and Inference #Statistics Theory (math.ST)
- Fast Rates for the Regret of Offline Reinforcement Learning
2021/01/31 by Yichun Hu, Hu, Yichun, Nathan Kallus +3 · 1 citation
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #Distributed Sensor Networks and Detection Algorithms #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Optimization and Control (math.OC) #Reinforcement Learning in Robotics
- Functional Graphical Models: Structure Enables Offline Data-Driven Optimization
2024/01/08 by Jakub Grudzien Kuba, Kuba, Jakub Grudzien, Masatoshi Uehara +5 · 1 citation
Computer Science · Materials Science · #Artificial Intelligence (cs.AI) #Computational Drug Discovery Methods #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Data Classification #Machine Learning in Materials Science