Ziyi Gao
- DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
DeepSeek-R1 shows an LLM can learn strong step-by-step reasoning from pure reinforcement learning, with no human-labeled reasoning examples.
2025/01/22 by DeepSeek-AI, Daya Guo, Dejian Yang +404 · 93 voices · 1715 citations
Computer Science · #Reinforcement Learning in Robotics #Data Stream Mining Techniques #Explainable Artificial Intelligence (XAI)
- DeepSeek-V3 Technical Report
2024/12/27 by DeepSeek-AI, Aixin Liu, Bei Feng +404 · 39 voices · 7 citations
Computer Science · Engineering · #Distributed and Parallel Computing Systems #Robotics and Automated Systems #cs.AI #cs.CL
- A novel laccase-like Cu-MOF for colorimetric differentiation and detection of phenolic compounds
2024/03/01 by Ziyi Gao, Jianping Guan, Meng Wang +4 · 2 citations
Materials Science · Engineering · #Advanced Nanomaterials in Catalysis #Electrochemical sensors and biosensors #Advanced Chemical Sensor Technologies