vix.ing · top · new · best · stats · spec

Zhibin Gou

  1. DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
    DeepSeek-R1 shows an LLM can learn strong step-by-step reasoning from pure reinforcement learning, with no human-labeled reasoning examples.
    2025/01/22 by DeepSeek-AI, Daya Guo, Guo, Daya +404 · 93 voices · 1649 citations
    Computer Science · #Reinforcement Learning in Robotics #Data Stream Mining Techniques #Explainable Artificial Intelligence (XAI)
  2. DeepSeek-V3 Technical Report
    2024/12/27 by DeepSeek-AI, Aixin Liu, Liu, Aixin +404 · 39 voices · 6 citations
    Computer Science · Engineering · #Distributed and Parallel Computing Systems #Robotics and Automated Systems #cs.AI #cs.CL
  3. CRITIC: Large Language Models Can Self-Correct with Tool-Interactive Critiquing
    2023/05/19 by Zhibin Gou, Zhihong Shao, Gou, Zhibin +11 · 1 voice · 108 citations
    Computer Science · #cs.CL #cs.AI
  4. DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence
    2024/06/17 by Qihao Zhu, DeepSeek-AI, Zhu, Qihao +79 · 1 voice · 69 citations
    Biochemistry, Genetics and Molecular Biology · Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Software Engineering (cs.SE) #cs.AI #cs.LG #cs.SE #vaccines and immunoinformatics approaches
  5. DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search
    2024/08/15 by Huajian Xin, Xin, Huajian, Z. Z. Ren +32 · 2 voices · 39 citations
    Computer Science · #Reinforcement Learning in Robotics #cs.AI #cs.CL #cs.LG #cs.LO
  6. ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving
    2023/09/29 by Zhibin Gou, Gou, Zhibin, Shao, Zhihong +12 · 45 citations
    Computer Science · Psychology · #Artificial Intelligence (cs.AI) #Artificial Intelligence in Games #Computation and Language (cs.CL) #Educational Games and Gamification #FOS: Computer and information sciences #Topic Modeling
  7. DeepSeek-Prover-V2: Advancing Formal Mathematical Reasoning via Reinforcement Learning for Subgoal Decomposition
    2025/04/30 by Z. Z. Ren, Ren, Z. Z., Zhihong Shao +33 · 4 voices · 57 citations
    #cs.CL #cs.AI
  8. Rho-1: Not All Tokens Are What You Need
    2024/04/11 by Zhenghao Lin, Zhibin Gou, Lin, Zhenghao +19 · 22 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Natural Language Processing Techniques #Speech Recognition and Synthesis #Topic Modeling
  9. Key-Point-Driven Data Synthesis with its Enhancement on Mathematical Reasoning
    2024/03/04 by Yiming Huang, Huang, Yiming, Xiao Liu +11 · 11 citations
    Computer Science · #Computability, Logic, AI Algorithms #Quantum Computing Algorithms and Architecture #Chaos-based Image/Signal Encryption
  10. CriticBench: Benchmarking LLMs for Critique-Correct Reasoning
    2024/02/22 by Zicheng Lin, Zhibin Gou, Lin, Zicheng +9 · 7 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Semantic Web and Ontologies
  11. Long Time No See! Open-Domain Conversation with Long-Term Persona Memory
    2022/03/11 by Xinchao Xu, Zhibin Gou, Xu, Xinchao +11 · 4 citations
    Computer Science · #AI in Service Interactions #Computation and Language (cs.CL) #FOS: Computer and information sciences #Persona Design and Applications #Topic Modeling
  12. DeepSeekMath-V2: Towards Self-Verifiable Mathematical Reasoning
    2025/11/27 by Zhihong Shao, Yuxiang Luo, Shao, Zhihong +15 · 7 citations
    Computer Science · Materials Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning in Materials Science #Mathematics, Computing, and Information Processing #Topic Modeling