vix.ing · top · new · best · stats · spec

Wang, Xiaoxiang

  1. DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
    DeepSeek-R1 shows an LLM can learn strong step-by-step reasoning from pure reinforcement learning, with no human-labeled reasoning examples.
    2025/01/22 by DeepSeek-AI, Daya Guo, Guo, Daya +404 · 93 voices · 2724 citations
    Computer Science · #Reinforcement Learning in Robotics #Data Stream Mining Techniques #Explainable Artificial Intelligence (XAI)
  2. DeepSeek-V3 Technical Report
    2024/12/27 by DeepSeek-AI, Aixin Liu, Liu, Aixin +404 · 39 voices · 7 citations
    Computer Science · Engineering · #Distributed and Parallel Computing Systems #Robotics and Automated Systems #cs.AI #cs.CL
  3. DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
    2024/05/07 by DeepSeek-AI, Aixin Liu, Bei Feng +310 · 5 voices · 368 citations
    Computer Science · #Expert finding and Q&A systems #Topic Modeling #Speech and dialogue systems
  4. DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models
    2025/12/02 by DeepSeek-AI, Aixin Liu, Aoxue Mei +352 · 46 citations
    Computer Science · Materials Science · Medicine · #Artificial Intelligence in Healthcare and Education #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning in Materials Science #Topic Modeling
  5. Panoramic Direct LiDAR-assisted Visual Odometry
    2024/09/14 by Hu, Qirui, Zikang Yuan, Yuan, Zikang +7 · 2 citations
    Engineering · Computer Science · Earth and Planetary Sciences · #Robotics and Sensor-Based Localization #Advanced Vision and Imaging #3D Surveying and Cultural Heritage
  6. LiDAR-Inertial Odometry in Dynamic Driving Scenarios using Label Consistency Detection
    2024/07/04 by Yuan, Zikang, Wang, Xiaoxiang, Wu, Jingying +2 · 1 citation
    #FOS: Computer and information sciences #Robotics (cs.RO)