Yuyang Zhou
- DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
DeepSeek-R1 shows an LLM can learn strong step-by-step reasoning from pure reinforcement learning, with no human-labeled reasoning examples.
2025/01/22 by DeepSeek-AI, Daya Guo, Dejian Yang +404 · 93 voices · 2643 citations
Computer Science · #Reinforcement Learning in Robotics #Data Stream Mining Techniques #Explainable Artificial Intelligence (XAI)
- DeepSeek-V3 Technical Report
2024/12/27 by DeepSeek-AI, Aixin Liu, Liu, Aixin +404 · 39 voices · 7 citations
Computer Science · Engineering · #Distributed and Parallel Computing Systems #Robotics and Automated Systems #cs.AI #cs.CL
- MoConVQ: Unified Physics-Based Motion Control via Scalable Discrete Representations
2023/10/16 by Heyuan Yao, Zhenhua Song, Yao, Heyuan +9 · 20 citations
Computer Science · Engineering · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Graphics (cs.GR) #Human Motion and Animation #Human Pose and Action Recognition #Multimodal Machine Learning Applications
- The Proximal Operator of the Piece-wise Exponential Function and Its Application in Compressed Sensing
2023/06/23 by Yulan Liu, Yuyang Zhou, Liu, Yulan +3 · 1 citation
Computer Science · Engineering · #FOS: Mathematics #Gaussian Processes and Bayesian Inference #Numerical Analysis (math.NA) #Optimization and Control (math.OC) #Sparse and Compressive Sensing Techniques #Sports Dynamics and Biomechanics