Shenzhi Wang
- Absolute Zero: Reinforced Self-play Reasoning with Zero Data
2025/05/06 by Andrew Zhao, Yiran Wu, Zhao, Andrew +20 · 23 voices · 76 citations
Computer Science · #Domain Adaptation and Few-Shot Learning #Multimodal Machine Learning Applications #Topic Modeling #cs.AI #cs.CL #cs.LG
- Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning
2025/06/02 by Shenzhi Wang, Le Yu, Wang, Shenzhi +37 · 1 voice · 173 citations
Computer Science · Social Sciences · #Artificial Intelligence in Law #Digital Rights Management and Security #Software Engineering Research #cs.AI #cs.CL #cs.LG
- DeeR-VLA: Dynamic Inference of Multimodal Large Language Models for Efficient Robot Execution
2024/11/04 by Yang Yue, Yue, Yang, Yulin Wang +13 · 35 citations
Computer Science · #AI-based Problem Solving and Planning #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Robotics (cs.RO) #Topic Modeling
- Avalon's Game of Thoughts: Battle Against Deception through Recursive Contemplation
2023/10/02 by Shenzhi Wang, Chang Liu, Wang, Shenzhi +17 · 7 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computers and Society (cs.CY) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multiagent Systems (cs.MA) #Natural Language Processing Techniques #Speech and dialogue systems #Topic Modeling
- Train Once, Get a Family: State-Adaptive Balances for Offline-to-Online Reinforcement Learning
2023/10/27 by Shenzhi Wang, Qisen Yang, Wang, Shenzhi +15 · 5 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Mobile Crowdsensing and Crowdsourcing #Reinforcement Learning in Robotics
- Hundreds Guide Millions: Adaptive Offline Reinforcement Learning with Expert Guidance
2023/09/04 by Qisen Yang, Shenzhi Wang, Yang, Qisen +7 · 4 citations
Computer Science · #Machine Learning and Data Classification #Reinforcement Learning in Robotics #Data Stream Mining Techniques
- Boosting Offline Reinforcement Learning with Action Preference Query
2023/06/06 by Qisen Yang, Shenzhi Wang, Yang, Qisen +7 · 3 citations
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Data Classification #Mobile Crowdsensing and Crowdsourcing #Reinforcement Learning in Robotics
- DiveR-CT: Diversity-enhanced Red Teaming Large Language Model Assistants with Relaxing Constraints
2024/05/29 by Andrew Zhao, Quentin Xu, Zhao, Andrew +11 · 2 citations
Medicine · #Radiomics and Machine Learning in Medical Imaging
- Scaffolded Language Models with Language Supervision for Mixed-Autonomy: A Survey
2024/10/21 by Matthieu Gaetan Lin, Jenny Sheng, Lin, Matthieu +16 · 1 citation
Engineering · Computer Science · #Industrial Technology and Control Systems #Fuzzy Logic and Control Systems #Fault Detection and Control Systems
- Kimi K3: Open Frontier Intelligence
2026/07/27 by Kimi Team, Tongtong Bai, Yifan Bai +398 · 1 voice
#cs.CL #cs.LG