vix.ing · top · new · best · stats · spec

Masatoshi Uehara

  1. Fine-Tuning of Continuous-Time Diffusion Models as Entropy-Regularized Control
    2024/02/23 by Masatoshi Uehara, Uehara, Masatoshi, Yulai Zhao +15 · 20 citations
    Computer Science · Mathematics · Physics and Astronomy · #Advanced Mathematical Modeling in Engineering #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Model Reduction and Neural Networks #Numerical methods in inverse problems
  2. A Review of Off-Policy Evaluation in Reinforcement Learning
    2022/12/13 by Masatoshi Uehara, Uehara, Masatoshi, Chengchun Shi +3 · 1 voice · 9 citations
    Computer Science · #Reinforcement Learning in Robotics #cs.LG #math.ST #stat.ME #stat.ML
  3. Understanding Reinforcement Learning-Based Fine-Tuning of Diffusion Models: A Tutorial and Review
    2024/07/18 by Masatoshi Uehara, Uehara, Masatoshi, Yulai Zhao +5 · 19 citations
    Engineering · #Artificial Intelligence (cs.AI) #FOS: Biological sciences #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Quantitative Methods (q-bio.QM) #Traffic control and management
  4. Double Reinforcement Learning for Efficient Off-Policy Evaluation in Markov Decision Processes
    2019/08/22 by Nathan Kallus, Masatoshi Uehara, Kallus, Nathan +1 · 11 citations
    Mathematics · Computer Science · Biochemistry, Genetics and Molecular Biology · #Advanced Causal Inference Techniques #Reinforcement Learning in Robotics #Gene Regulatory Network Analysis
  5. Inference-Time Alignment in Diffusion Models with Reward-Guided Generation: Tutorial and Review
    2025/01/16 by Masatoshi Uehara, Uehara, Masatoshi, Yulai Zhao +11 · 1 voice · 16 citations
    Physics and Astronomy · #Model Reduction and Neural Networks #cs.AI #cs.LG #q-bio.QM #stat.ML
  6. Minimax Weight and Q-Function Learning for Off-Policy Evaluation
    2019/10/28 by Masatoshi Uehara, Jiawei Huang, Uehara, Masatoshi +3 · 6 citations
    Computer Science · #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Formal Methods in Verification #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
  7. Fine-Tuning Discrete Diffusion Models via Reward Optimization with Applications to DNA and Protein Design
    2024/10/17 by Chenyu Wang, Masatoshi Uehara, Wang, Chenyu +17 · 15 citations
    Biochemistry, Genetics and Molecular Biology · #DNA and Nucleic Acid Chemistry
  8. Provable Offline Preference-Based Reinforcement Learning
    2023/05/24 by Wenhao Zhan, Masatoshi Uehara, Zhan, Wenhao +7 · 6 citations
    Biochemistry, Genetics and Molecular Biology · Computer Science · #Receptor Mechanisms and Signaling #Reinforcement Learning in Robotics #Formal Methods in Verification
  9. Pessimistic Model-based Offline Reinforcement Learning under Partial Coverage
    2021/07/13 by Masatoshi Uehara, Uehara, Masatoshi, W. Sun +1 · 4 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Reinforcement Learning in Robotics
  10. Feedback Efficient Online Fine-Tuning of Diffusion Models
    2024/02/26 by Masatoshi Uehara, Yulai Zhao, Uehara, Masatoshi +15 · 6 citations
    Mathematics · Computer Science · Physics and Astronomy · #Numerical methods for differential equations #Matrix Theory and Algorithms #Model Reduction and Neural Networks
  11. Localized Debiased Machine Learning: Efficient Inference on Quantile Treatment Effects and Beyond
    2019/12/30 by Nathan Kallus, Kallus, Nathan, Xiaojie Mao +3 · 3 citations
    Mathematics · #Advanced Causal Inference Techniques #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Methodology (stat.ME) #Statistical Methods and Bayesian Inference #Statistical Methods and Inference
  12. Dynamic Search for Inference-Time Alignment in Diffusion Models
    2025/03/03 by Xiner Li, Li, Xiner, Masatoshi Uehara +13 · 11 citations
    Physics and Astronomy · #Model Reduction and Neural Networks
  13. Provable Reward-Agnostic Preference-Based Reinforcement Learning
    2023/05/29 by Wenhao Zhan, Zhan, Wenhao, Masatoshi Uehara +5 · 4 citations
    Computer Science · #Advanced Graph Neural Networks #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Statistics Theory (math.ST)
  14. Representation Learning for Online and Offline RL in Low-rank MDPs
    2021/10/09 by Masatoshi Uehara, Uehara, Masatoshi, Xuezhou Zhang +3 · 3 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Reinforcement Learning in Robotics
  15. Provably Efficient Reinforcement Learning in Partially Observable Dynamical Systems
    2022/06/24 by Masatoshi Uehara, Ayush Sekhari, Uehara, Masatoshi +7 · 3 citations
    Computer Science · #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Methodology (stat.ME) #Reinforcement Learning in Robotics #Statistics Theory (math.ST)
  16. Finite Sample Analysis of Minimax Offline Reinforcement Learning: Completeness, Fast Rates and First-Order Efficiency
    2021/02/05 by Masatoshi Uehara, Masaaki Imaizumi, Uehara, Masatoshi +9 · 6 citations
    Computer Science · #FOS: Computer and information sciences #FOS: Mathematics #Formal Methods in Verification #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Software Reliability and Analysis Research #Statistics Theory (math.ST)
  17. Future-Dependent Value-Based Off-Policy Evaluation in POMDPs
    2022/07/26 by Masatoshi Uehara, Uehara, Masatoshi, Haruka Kiyohara +13 · 2 citations
    Engineering · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Smart Grid Energy Management
  18. Off-Policy Evaluation and Learning for External Validity under a\n Covariate Shift
    2020/02/26 by Masahiro Kato, Kato, Masahiro, Masatoshi Uehara +3 · 7 citations
    Mathematics · Engineering · Economics, Econometrics and Finance · #Advanced Causal Inference Techniques #Nuclear reactor physics and engineering #Economic Policies and Impacts
  19. Reward-Guided Iterative Refinement in Diffusion Models at Test-Time with Applications to Protein and DNA Design
    2025/02/20 by Masatoshi Uehara, Uehara, Masatoshi, Xingyu Su +13 · 1 voice · 3 citations
    Biochemistry, Genetics and Molecular Biology · Computer Science · Physics and Astronomy · #Gene Regulatory Network Analysis #Generative Adversarial Networks and Image Synthesis #Model Reduction and Neural Networks
  20. Mitigating Covariate Shift in Imitation Learning via Offline Data Without Great Coverage
    2021/06/06 by Jonathan Chang, Masatoshi Uehara, Chang, Jonathan D. +7 · 3 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Reinforcement Learning in Robotics
  21. Causal Inference Under Unmeasured Confounding With Negative Controls: A Minimax Learning Approach
    2021/03/25 by Nathan Kallus, Xiaojie Mao, Kallus, Nathan +3 · 1 citation
    Computer Science · Mathematics · #Advanced Causal Inference Techniques #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Methodology (stat.ME) #Statistical Methods and Inference
  22. Offline Minimax Soft-Q-learning Under Realizability and Partial Coverage
    2023/02/05 by Masatoshi Uehara, Uehara, Masatoshi, Nathan Kallus +5 · 2 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
  23. Minimax Instrumental Variable Regression and L2 Convergence Guarantees without Identification or Closedness
    2023/02/10 by Andrew F. Bennett, Nathan Kallus, Bennett, Andrew +9 · 1 citation
    Computer Science · Mathematics · #Distributed Sensor Networks and Detection Algorithms #Econometrics (econ.EM) #FOS: Computer and information sciences #FOS: Economics and business #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Methodology (stat.ME) #Statistical Methods and Inference #Statistics Theory (math.ST)
  24. Fast Rates for the Regret of Offline Reinforcement Learning
    2021/01/31 by Yichun Hu, Hu, Yichun, Nathan Kallus +3 · 1 citation
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #Distributed Sensor Networks and Detection Algorithms #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Optimization and Control (math.OC) #Reinforcement Learning in Robotics
  25. Functional Graphical Models: Structure Enables Offline Data-Driven Optimization
    2024/01/08 by Jakub Grudzien Kuba, Kuba, Jakub Grudzien, Masatoshi Uehara +5 · 1 citation
    Computer Science · Materials Science · #Artificial Intelligence (cs.AI) #Computational Drug Discovery Methods #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Data Classification #Machine Learning in Materials Science