vix.ing · top · new · best · stats · spec

Zhou, Ruida

  1. Natural Actor-Critic for Robust Reinforcement Learning with Function Approximation
    2023/07/17 by Ruida Zhou, Tao Liu, Zhou, Ruida +9 · 8 citations
    Computer Science · #Adaptive Dynamic Programming Control #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Optimization and Control (math.OC) #Reinforcement Learning in Robotics #Robotics (cs.RO)
  2. Learning Policies with Zero or Bounded Constraint Violation for\n Constrained MDPs
    2021/06/04 by Tao Liu, Ruida Zhou, Liu, Tao +7 · 2 citations
    Computer Science · #Reinforcement Learning in Robotics #Adversarial Robustness in Machine Learning #Formal Methods in Verification
  3. Anchor-Changing Regularized Natural Policy Gradient for Multi-Objective Reinforcement Learning
    2022/06/10 by Zhou, Ruida, Liu, Tao, Kalathil, Dileep +2 · 2 citations
    #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Optimization and Control (math.OC)
  4. Provably Fast Convergence of Independent Natural Policy Gradient for Markov Potential Games
    2023/10/15 by Sun, Youbang, Liu, Tao, Zhou, Ruida +2 · 2 citations
    #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Optimization and Control (math.OC)
  5. On the Learn-to-Optimize Capabilities of Transformers in In-Context Sparse Recovery
    2024/10/17 by Ruida Zhou, Liu, Renpu, Zhou, Ruida +4 · 3 citations
    Engineering · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Industrial Vision Systems and Defect Detection #Machine Learning (cs.LG)
  6. Regional Multi-Armed Bandits
    2018/02/22 by Wang, Zhiyang, Zhou, Ruida, Shen, Cong · 1 citation
    #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  7. Cost-aware Cascading Bandits
    2018/05/22 by Ruida Zhou, Zhou, Ruida, Chao Gan +5 · 1 citation
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Information Theory (cs.IT) #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Optimization and Search Problems
  8. Path-Guided Particle-based Sampling
    2024/12/04 by Mingzhou Fan, Ruida Zhou, Fan, Mingzhou +5 · 3 citations
    Biochemistry, Genetics and Molecular Biology · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Molecular Biology Techniques and Applications
  9. Individually Conditional Individual Mutual Information Bound on Generalization Error
    2020/12/17 by Zhou, Ruida, Tian, Chao, Liu, Tie · 1 citation
    #FOS: Computer and information sciences #Information Theory (cs.IT) #Machine Learning (stat.ML)
  10. Policy Optimization for Constrained MDPs with Provable Fast Global Convergence
    2021/10/31 by Tao Liu, Liu, Tao, Ruida Zhou +7 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning and Algorithms #Optimization and Control (math.OC) #Reinforcement Learning in Robotics #Stochastic Gradient Optimization Techniques
  11. Approximate Top-m Arm Identification with Heterogeneous Reward Variances
    2022/04/11 by Ruida Zhou, Chao Tian, Zhou, Ruida +1 · 1 citation
    Computer Science · Mathematics · #FOS: Computer and information sciences #Information Theory (cs.IT) #Machine Learning (cs.LG) #Machine Learning and Algorithms #Markov Chains and Monte Carlo Methods #Statistical Methods and Inference
  12. Data-adaptive Differentially Private Prompt Synthesis for In-Context Learning
    2024/10/15 by Ruida Zhou, Gao, Fengyu, Tianhao Wang +6 · 2 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Cryptography and Security (cs.CR) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Stochastic Gradient Optimization Techniques
  13. Provable Policy Gradient Methods for Average-Reward Markov Potential Games
    2024/03/09 by Min Cheng, Cheng, Min, Ruida Zhou +5 · 1 citation
    Computer Science · Decision Sciences · Engineering · #Advanced Control Systems Optimization #Computer Science and Game Theory (cs.GT) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #Risk and Portfolio Optimization
  14. Transformers learn variable-order Markov chains in-context
    2024/10/07 by Zhou, Ruida, Tian, Chao, Diggavi, Suhas · 1 citation
    #FOS: Computer and information sciences #Information Theory (cs.IT) #Machine Learning (cs.LG)
  15. On the Training Convergence of Transformers for In-Context Classification of Gaussian Mixtures
    2024/10/15 by Weiming Shen, Ruida Zhou, Shen, Wei +5 · 1 citation
    Computer Science · #Anomaly Detection Techniques and Applications #Context-Aware Activity Recognition Systems #FOS: Computer and information sciences #Information Theory (cs.IT) #Machine Learning (cs.LG) #Machine Learning (stat.ML)