vix.ing · top · new · best · stats · spec

Junxiao Song

  1. DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
    DeepSeek-R1 shows an LLM can learn strong step-by-step reasoning from pure reinforcement learning, with no human-labeled reasoning examples.
    2025/01/22 by DeepSeek-AI, Daya Guo, Guo, Daya +404 · 93 voices · 1882 citations
    Computer Science · #Reinforcement Learning in Robotics #Data Stream Mining Techniques #Explainable Artificial Intelligence (XAI)
  2. BioXP-0.5B: Explainable Medical-AI via RL-GRPO
    2024/02/05 by Zhihong Shao, Shao, Zhihong, Peiyi Wang +20 · 17 voices · 2210 citations
    Computer Science · #Mathematics, Computing, and Information Processing #Natural Language Processing Techniques #cs.AI #cs.CL #cs.LG
  3. DeepSeek-V3 Technical Report
    2024/12/27 by DeepSeek-AI, Aixin Liu, Bei Feng +404 · 39 voices · 7 citations
    Computer Science · Engineering · #Distributed and Parallel Computing Systems #Robotics and Automated Systems #cs.AI #cs.CL
  4. DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
    2024/05/07 by Aixin Liu, DeepSeek-AI, Liu, Aixin +310 · 5 voices · 264 citations
    Computer Science · #Expert finding and Q&A systems #Topic Modeling #Speech and dialogue systems
  5. DeepSeek LLM: Scaling Open-Source Language Models with Longtermism
    2024/01/05 by DeepSeek-AI, :, Xiao Bi +178 · 2 voices · 125 citations
    Computer Science · #Natural Language Processing Techniques #Text Readability and Simplification #Topic Modeling #cs.AI #cs.CL #cs.LG
  6. DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence
    2024/06/17 by DeepSeek-AI, Qihao Zhu, Daya Guo +79 · 1 voice · 86 citations
    Biochemistry, Genetics and Molecular Biology · Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Software Engineering (cs.SE) #cs.AI #cs.LG #cs.SE #vaccines and immunoinformatics approaches
  7. DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search
    2024/08/15 by Huajian Xin, Xin, Huajian, Z. Z. Ren +32 · 2 voices · 49 citations
    Computer Science · #Reinforcement Learning in Robotics #cs.AI #cs.CL #cs.LG #cs.LO
  8. DeepSeek-Prover-V2: Advancing Formal Mathematical Reasoning via Reinforcement Learning for Subgoal Decomposition
    2025/04/30 by Z. Z. Ren, Ren, Z. Z., Zhihong Shao +33 · 4 voices · 71 citations
    #cs.CL #cs.AI
  9. Sequence Design to Minimize the Weighted Integrated and Peak Sidelobe Levels
    2015/12/22 by Junxiao Song, Prabhu Babu, Daniel P. Palomar · 5 citations
    Engineering · Physics and Astronomy · #Antenna Design and Optimization #Radar Systems and Signal Processing #Radio Astronomy Observations and Technology