Yaohui Li
- DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
DeepSeek-R1 shows an LLM can learn strong step-by-step reasoning from pure reinforcement learning, with no human-labeled reasoning examples.
2025/01/22 by DeepSeek-AI, Daya Guo, Dejian Yang +404 · 93 voices · 2758 citations
Computer Science · #Reinforcement Learning in Robotics #Data Stream Mining Techniques #Explainable Artificial Intelligence (XAI)
- DeepSeek-V3 Technical Report
2024/12/27 by DeepSeek-AI, Aixin Liu, Bei Feng +404 · 39 voices · 7 citations
Computer Science · Engineering · #Distributed and Parallel Computing Systems #Robotics and Automated Systems #cs.AI #cs.CL
- DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
2024/05/07 by DeepSeek-AI, Aixin Liu, Bei Feng +310 · 5 voices · 376 citations
Computer Science · #Expert finding and Q&A systems #Topic Modeling #Speech and dialogue systems
- DiffUTE: Universal Text Editing Diffusion Model
2023/05/18 by Haoxing Chen, Zhuoer Xu, Chen, Haoxing +15 · 9 citations
Computer Science · #Generative Adversarial Networks and Image Synthesis #Computer Graphics and Visualization Techniques
- Multi-level Metric Learning for Few-shot Image Recognition
2021/03/21 by Haoxing Chen, Huaxiong Li, Chen, Haoxing +5 · 1 citation
Biochemistry, Genetics and Molecular Biology · Computer Science · #Cancer-related molecular mechanisms research #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Multimodal Machine Learning Applications
- The Effects of Sandstorms on the Climate of Northwestern China
2017/01/01 by Tiantian Hu, Di Wu, Yaohui Li +1 · 1 citation
Environmental Science · Earth and Planetary Sciences · #Atmospheric aerosols and clouds #Aeolian processes and effects #Meteorological Phenomena and Simulations