Zhou, Yuyang
- DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
DeepSeek-R1 shows an LLM can learn strong step-by-step reasoning from pure reinforcement learning, with no human-labeled reasoning examples.
2025/01/22 by DeepSeek-AI, Daya Guo, Guo, Daya +404 · 93 voices · 2633 citations
Computer Science · #Reinforcement Learning in Robotics #Data Stream Mining Techniques #Explainable Artificial Intelligence (XAI)
- DeepSeek-V3 Technical Report
2024/12/27 by DeepSeek-AI, Aixin Liu, Liu, Aixin +404 · 39 voices · 7 citations
Computer Science · Engineering · #Distributed and Parallel Computing Systems #Robotics and Automated Systems #cs.AI #cs.CL
- MoConVQ: Unified Physics-Based Motion Control via Scalable Discrete Representations
2023/10/16 by Heyuan Yao, Yao, Heyuan, Zhenhua Song +9 · 19 citations
Computer Science · Engineering · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Graphics (cs.GR) #Human Motion and Animation #Human Pose and Action Recognition #Multimodal Machine Learning Applications
- DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models
2025/12/02 by DeepSeek-AI, Aixin Liu, Aoxue Mei +352 · 46 citations
Computer Science · Materials Science · Medicine · #Artificial Intelligence in Healthcare and Education #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning in Materials Science #Topic Modeling
- The Proximal Operator of the Piece-wise Exponential Function and Its Application in Compressed Sensing
2023/06/23 by Yulan Liu, Yuyang Zhou, Liu, Yulan +3 · 1 citation
Computer Science · Engineering · #FOS: Mathematics #Gaussian Processes and Bayesian Inference #Numerical Analysis (math.NA) #Optimization and Control (math.OC) #Sparse and Compressive Sensing Techniques #Sports Dynamics and Biomechanics