Wang, Shenzhi
- Absolute Zero: Reinforced Self-play Reasoning with Zero Data
2025/05/06 by Andrew Zhao, Yilong Wu, Yiran Wu +20 · 23 voices · 76 citations
Computer Science · #Domain Adaptation and Few-Shot Learning #Multimodal Machine Learning Applications #Topic Modeling #cs.AI #cs.CL #cs.LG
- Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning
2025/06/02 by Shenzhi Wang, Wang, Shenzhi, Le Yu +37 · 1 voice · 173 citations
Computer Science · Social Sciences · #Artificial Intelligence in Law #Digital Rights Management and Security #Software Engineering Research #cs.AI #cs.CL #cs.LG
- DeeR-VLA: Dynamic Inference of Multimodal Large Language Models for Efficient Robot Execution
2024/11/04 by Yang Yue, Yulin Wang, Yue, Yang +13 · 35 citations
Computer Science · #AI-based Problem Solving and Planning #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Robotics (cs.RO) #Topic Modeling
- Avalon's Game of Thoughts: Battle Against Deception through Recursive Contemplation
2023/10/02 by Shenzhi Wang, Chang Liu, Wang, Shenzhi +17 · 7 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computers and Society (cs.CY) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multiagent Systems (cs.MA) #Natural Language Processing Techniques #Speech and dialogue systems #Topic Modeling
- Train Once, Get a Family: State-Adaptive Balances for Offline-to-Online Reinforcement Learning
2023/10/27 by Shenzhi Wang, Wang, Shenzhi, Qisen Yang +15 · 5 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Mobile Crowdsensing and Crowdsourcing #Reinforcement Learning in Robotics
- Hundreds Guide Millions: Adaptive Offline Reinforcement Learning with Expert Guidance
2023/09/04 by Qisen Yang, Yang, Qisen, Shenzhi Wang +7 · 4 citations
Computer Science · #Machine Learning and Data Classification #Reinforcement Learning in Robotics #Data Stream Mining Techniques
- PsychoGAT: A Novel Psychological Measurement Paradigm through Interactive Fiction Games with LLM Agents
2024/02/19 by Yang, Qisen, Wang, Zekun, Chen, Honghui +6 · 4 citations
#Computation and Language (cs.CL) #Computers and Society (cs.CY) #FOS: Computer and information sciences #Human-Computer Interaction (cs.HC) #Machine Learning (cs.LG) #Multiagent Systems (cs.MA)
- Boosting Offline Reinforcement Learning with Action Preference Query
2023/06/06 by Qisen Yang, Yang, Qisen, Shenzhi Wang +7 · 3 citations
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Data Classification #Mobile Crowdsensing and Crowdsourcing #Reinforcement Learning in Robotics
- Model Surgery: Modulating LLM's Behavior Via Simple Parameter Editing
2024/07/11 by Wang, Huanqian, Yue, Yang, Lu, Rui +5 · 4 citations
#62M45 (Secondary) #68T50 (Primary) 68T07 #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #I.2.7
- OS Agents: A Survey on MLLM-based Agents for General Computing Devices Use
2025/08/06 by Hu, Xueyu, Xiong, Tao, Yi, Biao +26 · 16 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- DiveR-CT: Diversity-enhanced Red Teaming Large Language Model Assistants with Relaxing Constraints
2024/05/29 by Andrew Zhao, Zhao, Andrew, Quentin Xu +11 · 2 citations
Medicine · #Radiomics and Machine Learning in Medical Imaging
- Scaffolded Language Models with Language Supervision for Mixed-Autonomy: A Survey
2024/10/21 by Matthieu Gaetan Lin, Lin, Matthieu, Jenny Sheng +16 · 1 citation
Engineering · Computer Science · #Industrial Technology and Control Systems #Fuzzy Logic and Control Systems #Fault Detection and Control Systems