vix.ing · top · new · best · stats · spec

Luo, Fuli

  1. DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
    DeepSeek-R1 shows an LLM can learn strong step-by-step reasoning from pure reinforcement learning, with no human-labeled reasoning examples.
    2025/01/22 by DeepSeek-AI, Daya Guo, Guo, Daya +404 · 93 voices · 1894 citations
    Computer Science · #Reinforcement Learning in Robotics #Data Stream Mining Techniques #Explainable Artificial Intelligence (XAI)
  2. DeepSeek-V3 Technical Report
    2024/12/27 by DeepSeek-AI, Aixin Liu, Bei Feng +404 · 39 voices · 7 citations
    Computer Science · Engineering · #Distributed and Parallel Computing Systems #Robotics and Automated Systems #cs.AI #cs.CL
  3. DeepSeekMoE: Towards Ultimate Expert Specialization in Mixture-of-Experts Language Models
    2024/01/11 by Damai Dai, Chengqi Deng, Dai, Damai +33 · 5 voices · 201 citations
    Computer Science · #Computation and Language (cs.CL) #Expert finding and Q&A systems #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Topic Modeling #cs.CL
  4. DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
    2024/05/07 by DeepSeek-AI, Aixin Liu, Bei Feng +310 · 5 voices · 267 citations
    Computer Science · #Expert finding and Q&A systems #Topic Modeling #Speech and dialogue systems
  5. DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence
    2024/01/25 by Daya Guo, Guo, Daya, Qihao Zhu +25 · 3 voices · 289 citations
    Computer Science · #Natural Language Processing Techniques #Software Engineering Research #Topic Modeling #cs.CL #cs.LG #cs.SE
  6. DeepSeek LLM: Scaling Open-Source Language Models with Longtermism
    2024/01/05 by DeepSeek-AI, Xiao Guo Bi, : +178 · 2 voices · 126 citations
    Computer Science · #Natural Language Processing Techniques #Text Readability and Simplification #Topic Modeling #cs.AI #cs.CL #cs.LG
  7. DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence
    2024/06/17 by DeepSeek-AI, Qihao Zhu, Daya Guo +79 · 1 voice · 86 citations
    Biochemistry, Genetics and Molecular Biology · Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Software Engineering (cs.SE) #cs.AI #cs.LG #cs.SE #vaccines and immunoinformatics approaches
  8. DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search
    2024/08/15 by Huajian Xin, Z. Z. Ren, Xin, Huajian +32 · 2 voices · 50 citations
    Computer Science · #Reinforcement Learning in Robotics #cs.AI #cs.CL #cs.LG #cs.LO
  9. Raise a Child in Large Language Model: Towards Effective and Generalizable Fine-tuning
    2021/09/13 by Runxin Xu, Fuli Luo, Xu, Runxin +11 · 11 citations
    Computer Science · #Topic Modeling #Domain Adaptation and Few-Shot Learning #Speech Recognition and Synthesis
  10. Pun-GAN: Generative Adversarial Network for Pun Generation
    2019/10/24 by Fuli Luo, Luo, Fuli, Shunyao Li +11 · 3 citations
    Computer Science · Psychology · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Humor Studies and Applications #Multimodal Machine Learning Applications
  11. Towards Unified Prompt Tuning for Few-shot Text Classification
    2022/05/11 by Jianing Wang, Chengyu Wang, Wang, Jianing +15 · 3 citations
    Computer Science · #Topic Modeling #Natural Language Processing Techniques #Multimodal Machine Learning Applications
  12. From Dense to Sparse: Contrastive Pruning for Better Pre-trained Language Model Compression
    2021/12/14 by Xu, Runxin, Luo, Fuli, Wang, Chengyu +4 · 2 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
  13. Parameter-Efficient Sparsity for Large Language Models Fine-Tuning
    2022/05/23 by Yuchao Li, Fuli Luo, Li, Yuchao +11 · 2 citations
    Computer Science · #Topic Modeling #Natural Language Processing Techniques #Multimodal Machine Learning Applications
  14. A Dual Reinforcement Learning Framework for Unsupervised Text Style Transfer
    2019/05/24 by Luo, Fuli, Li, Peng, Zhou, Jie +4 · 1 citation
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  15. VECO: Variable and Flexible Cross-lingual Pre-training for Language Understanding and Generation
    2020/10/30 by Luo, Fuli, Wang, Wei, Liu, Jiahao +5 · 1 citation
    #Computation and Language (cs.CL) #FOS: Computer and information sciences
  16. Making Pre-trained Language Models End-to-end Few-shot Learners with Contrastive Prompt Tuning
    2022/04/01 by Xu, Ziyun, Wang, Chengyu, Qiu, Minghui +4 · 1 citation
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences
  17. Stabilizing MoE Reinforcement Learning by Aligning Training and Inference Routers
    2025/10/13 by Wujun Ma, Ma, Wenhan, Hailin Zhang +11 · 5 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #Topic Modeling
  18. MiMo-Embodied: X-Embodied Foundation Model Technical Report
    2025/11/20 by Xiaoshuai Hao, Lei Zhou, Hao, Xiaoshuai +76 · 3 citations
    Engineering · Computer Science · Psychology · #Autonomous Vehicle Technology and Safety #Multimodal Machine Learning Applications #Social Robot Interaction and HRI