Chengshuai Shi
- Efficient Prompt Optimization Through the Lens of Best Arm Identification
2024/02/15 by Chengshuai Shi, Kun Yang, Shi, Chengshuai +5 · 8 citations
Computer Science · Engineering · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Intelligent Tutoring Systems and Adaptive Learning #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Robot Manipulation and Learning
- Continual Harness: Online Adaptation for Self-Improving Foundation Agents
2026/05/11 by Seth Karten, Joel Zhang, Tersoo Upaa +5 · 2 voices · 1 citation
#cs.LG #cs.AI
- Heterogeneous Multi-player Multi-armed Bandits: Closing the Gap and Generalization
2021/10/27 by Chengshuai Shi, Shi, Chengshuai, Wei Xiong +5 · 3 citations
Computer Science · Decision Sciences · Engineering · #Advanced Bandit Algorithms Research #Cognitive Radio Networks and Spectrum Sensing #FOS: Computer and information sciences #Information Theory (cs.IT) #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Smart Grid Energy Management
- Decentralized Multi-player Multi-armed Bandits with No Collision Information
2020/02/29 by Chengshuai Shi, Wei Xiong, Shi, Chengshuai +5 · 2 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Auction Theory and Applications #FOS: Computer and information sciences #Information Theory (cs.IT) #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Optimization and Search Problems
- Nearly Minimax Optimal Offline Reinforcement Learning with Linear Function Approximation: Single-Agent MDP and Markov Game
2022/05/31 by Wei Xiong, Han Zhong, Xiong, Wei +9 · 2 citations
Computer Science · #Computer Science and Game Theory (cs.GT) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
- On High-dimensional and Low-rank Tensor Bandits
2023/05/06 by Chengshuai Shi, Cong Shen, Shi, Chengshuai +3 · 1 citation
Decision Sciences · Mathematics · Engineering · #Advanced Bandit Algorithms Research #Tensor decomposition and applications #Energy Load and Power Forecasting
- LeAct: Learning to Reason from Expert Actions
2026/07/23 by Ziran Yang, Chengshuai Shi, Raj Ghugare +3
#cs.LG #cs.AI #cs.CL