Jiali Cai
- DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
DeepSeek-R1 shows an LLM can learn strong step-by-step reasoning from pure reinforcement learning, with no human-labeled reasoning examples.
2025/01/22 by DeepSeek-AI, Daya Guo, Dejian Yang +404 · 93 voices · 2750 citations
Computer Science · #Reinforcement Learning in Robotics #Data Stream Mining Techniques #Explainable Artificial Intelligence (XAI)
- DeepSeek-V3 Technical Report
2024/12/27 by DeepSeek-AI, Aixin Liu, Bei Feng +404 · 39 voices · 7 citations
Computer Science · Engineering · #Distributed and Parallel Computing Systems #Robotics and Automated Systems #cs.AI #cs.CL
- Magneto-optical trapping of aluminum monofluoride
2025/06/02 by J. E. Padilla‐Castillo, Padilla-Castillo, J. E., J. Cai +18 · 1 voice · 7 citations
Computer Science · Physics and Astronomy · #Advanced Fiber Laser Technologies #Atomic Physics (physics.atom-ph) #Cold Atom Physics and Bose-Einstein Condensates #FOS: Physical sciences #Quantum Physics (quant-ph) #Quantum optics and atomic interactions