vix.ing · top · new · best · stats · spec

Jiaqi Ni

  1. DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
    DeepSeek-R1 shows an LLM can learn strong step-by-step reasoning from pure reinforcement learning, with no human-labeled reasoning examples.
    2025/01/22 by DeepSeek-AI, Daya Guo, Dejian Yang +404 · 93 voices · 2740 citations
    Computer Science · #Reinforcement Learning in Robotics #Data Stream Mining Techniques #Explainable Artificial Intelligence (XAI)
  2. DeepSeek-V3 Technical Report
    2024/12/27 by DeepSeek-AI, Aixin Liu, Bei Feng +404 · 39 voices · 7 citations
    Computer Science · Engineering · #Distributed and Parallel Computing Systems #Robotics and Automated Systems #cs.AI #cs.CL
  3. DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
    2024/05/07 by Aixin Liu, DeepSeek-AI, Liu, Aixin +310 · 5 voices · 373 citations
    Computer Science · #Expert finding and Q&A systems #Topic Modeling #Speech and dialogue systems
  4. Invariant subspaces of weighted Bergman spaces in infinitely many variables
    2021/03/06 by Hui Dan, Kunyu Guo, Dan, Hui +3 · 1 citation
    Mathematics · #Advanced Harmonic Analysis Research #Algebraic and Geometric Analysis #Complex Variables (math.CV) #FOS: Mathematics #Functional Analysis (math.FA) #Holomorphic and Operator Theory