Liu, Aixin
- DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
DeepSeek-R1 shows an LLM can learn strong step-by-step reasoning from pure reinforcement learning, with no human-labeled reasoning examples.
2025/01/22 by DeepSeek-AI, Daya Guo, Guo, Daya +404 · 93 voices · 2113 citations
Computer Science · #Reinforcement Learning in Robotics #Data Stream Mining Techniques #Explainable Artificial Intelligence (XAI)
- DeepSeek-V3 Technical Report
2024/12/27 by DeepSeek-AI, Aixin Liu, Bei Feng +404 · 39 voices · 7 citations
Computer Science · Engineering · #Distributed and Parallel Computing Systems #Robotics and Automated Systems #cs.AI #cs.CL
- DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
2024/05/07 by Aixin Liu, DeepSeek-AI, Liu, Aixin +310 · 5 voices · 298 citations
Computer Science · #Expert finding and Q&A systems #Topic Modeling #Speech and dialogue systems
- DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding
2024/12/13 by Zhiyu Wu, Xiaokang Chen, Wu, Zhiyu +51 · 3 voices · 185 citations
Computer Science · #cs.CV #cs.AI #cs.CL
- DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence
2024/06/17 by DeepSeek-AI, Qihao Zhu, Zhu, Qihao +79 · 1 voice · 97 citations
Biochemistry, Genetics and Molecular Biology · Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Software Engineering (cs.SE) #cs.AI #cs.LG #cs.SE #vaccines and immunoinformatics approaches
- DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models
2025/12/02 by DeepSeek-AI, Liu, Aixin, Mei, Aoxue +260 · 46 citations
#Computation and Language (cs.CL) #FOS: Computer and information sciences