vix.ing · top · new · best · stats · spec

Guanting Chen

  1. DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
    DeepSeek-R1 shows an LLM can learn strong step-by-step reasoning from pure reinforcement learning, with no human-labeled reasoning examples.
    2025/01/22 by DeepSeek-AI, Daya Guo, Guo, Daya +404 · 93 voices · 2286 citations
    Computer Science · #Reinforcement Learning in Robotics #Data Stream Mining Techniques #Explainable Artificial Intelligence (XAI)
  2. DeepSeek-V3 Technical Report
    2024/12/27 by DeepSeek-AI, Aixin Liu, Bei Feng +404 · 39 voices · 7 citations
    Computer Science · Engineering · #Distributed and Parallel Computing Systems #Robotics and Automated Systems #cs.AI #cs.CL
  3. DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
    2024/05/07 by DeepSeek-AI, Aixin Liu, Bei Feng +310 · 5 voices · 319 citations
    Computer Science · #Expert finding and Q&A systems #Topic Modeling #Speech and dialogue systems
  4. DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence
    2024/01/25 by Daya Guo, Qihao Zhu, Guo, Daya +25 · 3 voices · 341 citations
    Computer Science · #Natural Language Processing Techniques #Software Engineering Research #Topic Modeling #cs.CL #cs.LG #cs.SE
  5. DeepSeek LLM: Scaling Open-Source Language Models with Longtermism
    2024/01/05 by DeepSeek-AI, :, Xiao Bi +178 · 2 voices · 163 citations
    Computer Science · #Natural Language Processing Techniques #Text Readability and Simplification #Topic Modeling #cs.AI #cs.CL #cs.LG
  6. Fire-Flyer AI-HPC: A Cost-Effective Software-Hardware Co-Design for Deep Learning
    2024/08/26 by Wei An, Xiao Bi, An, Wei +107 · 2 voices · 5 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Distributed #FOS: Computer and information sciences #Parallel #Parallel Computing and Optimization Techniques #and Cluster Computing (cs.DC) #cs.AI #cs.DC
  7. An Adaptive State Aggregation Algorithm for Markov Decision Processes
    2021/07/23 by Guanting Chen, Johann Demetrio Gaebler, Chen, Guanting +7 · 1 citation
    Computer Science · #Bayesian Modeling and Causal Inference #Data Stream Mining Techniques #Data Structures and Algorithms (cs.DS) #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Optimization and Control (math.OC) #Reinforcement Learning in Robotics
  8. Towards Better Understanding of In-Context Learning Ability from In-Context Uncertainty Quantification
    2024/05/24 by Shang Liu, Zhongze Cai, Liu, Shang +5 · 1 citation
    Computer Science · Engineering · #AI-based Problem Solving and Planning #Computation and Language (cs.CL) #FOS: Computer and information sciences #Fault Detection and Control Systems #Intelligent Tutoring Systems and Adaptive Learning #Machine Learning (cs.LG) #Machine Learning (stat.ML)
  9. Reward Modeling with Ordinal Feedback: Wisdom of the Crowd
    2024/11/19 by Shang Liu, Liu, Shang, Yu Pan +5 · 1 citation
    Economics, Econometrics and Finance · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Diverse Scientific and Economic Studies #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)