vix.ing · top · new · best · stats · spec

Shenzhi Wang

  1. Absolute Zero: Reinforced Self-play Reasoning with Zero Data
    2025/05/06 by Andrew Zhao, Yiran Wu, Zhao, Andrew +20 · 23 voices · 76 citations
    Computer Science · #Domain Adaptation and Few-Shot Learning #Multimodal Machine Learning Applications #Topic Modeling #cs.AI #cs.CL #cs.LG
  2. Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning
    2025/06/02 by Shenzhi Wang, Le Yu, Wang, Shenzhi +37 · 1 voice · 173 citations
    Computer Science · Social Sciences · #Artificial Intelligence in Law #Digital Rights Management and Security #Software Engineering Research #cs.AI #cs.CL #cs.LG
  3. DeeR-VLA: Dynamic Inference of Multimodal Large Language Models for Efficient Robot Execution
    2024/11/04 by Yang Yue, Yue, Yang, Yulin Wang +13 · 35 citations
    Computer Science · #AI-based Problem Solving and Planning #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Robotics (cs.RO) #Topic Modeling
  4. Avalon's Game of Thoughts: Battle Against Deception through Recursive Contemplation
    2023/10/02 by Shenzhi Wang, Chang Liu, Wang, Shenzhi +17 · 7 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computers and Society (cs.CY) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multiagent Systems (cs.MA) #Natural Language Processing Techniques #Speech and dialogue systems #Topic Modeling
  5. Train Once, Get a Family: State-Adaptive Balances for Offline-to-Online Reinforcement Learning
    2023/10/27 by Shenzhi Wang, Qisen Yang, Wang, Shenzhi +15 · 5 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Mobile Crowdsensing and Crowdsourcing #Reinforcement Learning in Robotics
  6. Hundreds Guide Millions: Adaptive Offline Reinforcement Learning with Expert Guidance
    2023/09/04 by Qisen Yang, Shenzhi Wang, Yang, Qisen +7 · 4 citations
    Computer Science · #Machine Learning and Data Classification #Reinforcement Learning in Robotics #Data Stream Mining Techniques
  7. Boosting Offline Reinforcement Learning with Action Preference Query
    2023/06/06 by Qisen Yang, Shenzhi Wang, Yang, Qisen +7 · 3 citations
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Data Classification #Mobile Crowdsensing and Crowdsourcing #Reinforcement Learning in Robotics
  8. DiveR-CT: Diversity-enhanced Red Teaming Large Language Model Assistants with Relaxing Constraints
    2024/05/29 by Andrew Zhao, Quentin Xu, Zhao, Andrew +11 · 2 citations
    Medicine · #Radiomics and Machine Learning in Medical Imaging
  9. Scaffolded Language Models with Language Supervision for Mixed-Autonomy: A Survey
    2024/10/21 by Matthieu Gaetan Lin, Jenny Sheng, Lin, Matthieu +16 · 1 citation
    Engineering · Computer Science · #Industrial Technology and Control Systems #Fuzzy Logic and Control Systems #Fault Detection and Control Systems
  10. Kimi K3: Open Frontier Intelligence
    2026/07/27 by Kimi Team, Tongtong Bai, Yifan Bai +398 · 1 voice
    #cs.CL #cs.LG