Zhou, Ruida
- Natural Actor-Critic for Robust Reinforcement Learning with Function Approximation
2023/07/17 by Ruida Zhou, Tao Liu, Zhou, Ruida +9 · 8 citations
Computer Science · #Adaptive Dynamic Programming Control #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Optimization and Control (math.OC) #Reinforcement Learning in Robotics #Robotics (cs.RO)
- Learning Policies with Zero or Bounded Constraint Violation for\n Constrained MDPs
2021/06/04 by Tao Liu, Ruida Zhou, Liu, Tao +7 · 2 citations
Computer Science · #Reinforcement Learning in Robotics #Adversarial Robustness in Machine Learning #Formal Methods in Verification
- Anchor-Changing Regularized Natural Policy Gradient for Multi-Objective Reinforcement Learning
2022/06/10 by Zhou, Ruida, Liu, Tao, Kalathil, Dileep +2 · 2 citations
#FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Optimization and Control (math.OC)
- Provably Fast Convergence of Independent Natural Policy Gradient for Markov Potential Games
2023/10/15 by Sun, Youbang, Liu, Tao, Zhou, Ruida +2 · 2 citations
#FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Optimization and Control (math.OC)
- On the Learn-to-Optimize Capabilities of Transformers in In-Context Sparse Recovery
2024/10/17 by Ruida Zhou, Liu, Renpu, Zhou, Ruida +4 · 3 citations
Engineering · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Industrial Vision Systems and Defect Detection #Machine Learning (cs.LG)
- Regional Multi-Armed Bandits
2018/02/22 by Wang, Zhiyang, Zhou, Ruida, Shen, Cong · 1 citation
#FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Cost-aware Cascading Bandits
2018/05/22 by Ruida Zhou, Zhou, Ruida, Chao Gan +5 · 1 citation
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Information Theory (cs.IT) #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Optimization and Search Problems
- Path-Guided Particle-based Sampling
2024/12/04 by Mingzhou Fan, Ruida Zhou, Fan, Mingzhou +5 · 3 citations
Biochemistry, Genetics and Molecular Biology · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Molecular Biology Techniques and Applications
- Individually Conditional Individual Mutual Information Bound on Generalization Error
2020/12/17 by Zhou, Ruida, Tian, Chao, Liu, Tie · 1 citation
#FOS: Computer and information sciences #Information Theory (cs.IT) #Machine Learning (stat.ML)
- Policy Optimization for Constrained MDPs with Provable Fast Global Convergence
2021/10/31 by Tao Liu, Liu, Tao, Ruida Zhou +7 · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning and Algorithms #Optimization and Control (math.OC) #Reinforcement Learning in Robotics #Stochastic Gradient Optimization Techniques
- Approximate Top-m Arm Identification with Heterogeneous Reward Variances
2022/04/11 by Ruida Zhou, Chao Tian, Zhou, Ruida +1 · 1 citation
Computer Science · Mathematics · #FOS: Computer and information sciences #Information Theory (cs.IT) #Machine Learning (cs.LG) #Machine Learning and Algorithms #Markov Chains and Monte Carlo Methods #Statistical Methods and Inference
- Data-adaptive Differentially Private Prompt Synthesis for In-Context Learning
2024/10/15 by Ruida Zhou, Gao, Fengyu, Tianhao Wang +6 · 2 citations
Computer Science · #Artificial Intelligence (cs.AI) #Cryptography and Security (cs.CR) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Stochastic Gradient Optimization Techniques
- Provable Policy Gradient Methods for Average-Reward Markov Potential Games
2024/03/09 by Min Cheng, Cheng, Min, Ruida Zhou +5 · 1 citation
Computer Science · Decision Sciences · Engineering · #Advanced Control Systems Optimization #Computer Science and Game Theory (cs.GT) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #Risk and Portfolio Optimization
- Transformers learn variable-order Markov chains in-context
2024/10/07 by Zhou, Ruida, Tian, Chao, Diggavi, Suhas · 1 citation
#FOS: Computer and information sciences #Information Theory (cs.IT) #Machine Learning (cs.LG)
- On the Training Convergence of Transformers for In-Context Classification of Gaussian Mixtures
2024/10/15 by Weiming Shen, Ruida Zhou, Shen, Wei +5 · 1 citation
Computer Science · #Anomaly Detection Techniques and Applications #Context-Aware Activity Recognition Systems #FOS: Computer and information sciences #Information Theory (cs.IT) #Machine Learning (cs.LG) #Machine Learning (stat.ML)