vix.ing · top · new · best · stats · spec

Yu, Kuai

  1. DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
    DeepSeek-R1 shows an LLM can learn strong step-by-step reasoning from pure reinforcement learning, with no human-labeled reasoning examples.
    2025/01/22 by DeepSeek-AI, Daya Guo, Guo, Daya +404 · 93 voices · 1702 citations
    Computer Science · #Reinforcement Learning in Robotics #Data Stream Mining Techniques #Explainable Artificial Intelligence (XAI)
  2. DeepSeek-V3 Technical Report
    2024/12/27 by DeepSeek-AI, Aixin Liu, Liu, Aixin +404 · 39 voices · 7 citations
    Computer Science · Engineering · #Distributed and Parallel Computing Systems #Robotics and Automated Systems #cs.AI #cs.CL
  3. DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models
    2025/12/02 by DeepSeek-AI, Liu, Aixin, Mei, Aoxue +260 · 44 citations
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  4. Insights into LLM Long-Context Failures: When Transformers Know but Don't Tell
    2024/06/20 by Taiming Lu, Muhan Gao, Lu, Taiming +7 · 7 citations
    Engineering · #Advancements in Photolithography Techniques #Electricity Theft Detection Techniques
  5. Large Processor Chip Model
    2025/06/03 by Chang, Kaiyan, Chen, Mingzhi, Chen, Yunji +40 · 2 citations
    #FOS: Computer and information sciences #Hardware Architecture (cs.AR)