Gao, Ziyi
- DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
DeepSeek-R1 shows an LLM can learn strong step-by-step reasoning from pure reinforcement learning, with no human-labeled reasoning examples.
2025/01/22 by DeepSeek-AI, Daya Guo, Dejian Yang +404 · 93 voices · 1985 citations
Computer Science · #Reinforcement Learning in Robotics #Data Stream Mining Techniques #Explainable Artificial Intelligence (XAI)
- DeepSeek-V3 Technical Report
2024/12/27 by DeepSeek-AI, Aixin Liu, Liu, Aixin +404 · 39 voices · 7 citations
Computer Science · Engineering · #Distributed and Parallel Computing Systems #Robotics and Automated Systems #cs.AI #cs.CL
- DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models
2025/12/02 by DeepSeek-AI, Liu, Aixin, Mei, Aoxue +260 · 44 citations
#Computation and Language (cs.CL) #FOS: Computer and information sciences
- ReToMe-VA: Recursive Token Merging for Video Diffusion-based Unrestricted Adversarial Attack
2024/08/10 by Ziyi Gao, Kai Chen, Gao, Ziyi +13 · 1 citation
Computer Science · #Adversarial Robustness in Machine Learning #Computer Vision and Pattern Recognition (cs.CV) #Digital Media Forensic Detection #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis