vix.ing · top · new · best · stats · spec

Xiaokang Chen

  1. DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
    DeepSeek-R1 shows an LLM can learn strong step-by-step reasoning from pure reinforcement learning, with no human-labeled reasoning examples.
    2025/01/22 by DeepSeek-AI, Daya Guo, Dejian Yang +404 · 93 voices · 1668 citations
    Computer Science · #Reinforcement Learning in Robotics #Data Stream Mining Techniques #Explainable Artificial Intelligence (XAI)
  2. DeepSeek-V3 Technical Report
    2024/12/27 by DeepSeek-AI, Aixin Liu, Bei Feng +404 · 39 voices · 6 citations
    Computer Science · Engineering · #Distributed and Parallel Computing Systems #Robotics and Automated Systems #cs.AI #cs.CL
  3. DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
    2024/05/07 by Aixin Liu, DeepSeek-AI, Bei Feng +310 · 5 voices · 237 citations
    Computer Science · #Expert finding and Q&A systems #Topic Modeling #Speech and dialogue systems
  4. DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding
    2024/12/13 by Zhiyu Wu, Wu, Zhiyu, Xiaokang Chen +51 · 3 voices · 140 citations
    Computer Science · #cs.CV #cs.AI #cs.CL
  5. Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling
    2025/01/29 by Xiaokang Chen, Zhiyu Wu, Chen, Xiaokang +13 · 1 voice · 227 citations
    Computer Science · #Semantic Web and Ontologies #cs.AI #cs.CL #cs.CV
  6. LGM: Large Multi-View Gaussian Model for High-Resolution 3D Content Creation
    2024/02/07 by Jiaxiang Tang, Tang, Jiaxiang, Zhaoxi Chen +9 · 138 citations
    Computer Science · Engineering · #3D Shape Modeling and Analysis #Computer Graphics and Visualization Techniques #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Image Processing and 3D Reconstruction
  7. Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation
    2024/10/17 by Chengyue Wu, Wu, Chengyue, Xiaokang Chen +19 · 113 citations
    Computer Science · #Multimodal Machine Learning Applications
  8. VisionLLM: Large Language Model is also an Open-Ended Decoder for Vision-Centric Tasks
    2023/05/18 by Wenhai Wang, Wang, Wenhai, Zhe Chen +19 · 55 citations
    Computer Science · #Multimodal Machine Learning Applications #Domain Adaptation and Few-Shot Learning #Natural Language Processing Techniques
  9. Semi-Supervised Semantic Segmentation with Cross Pseudo Supervision
    2021/06/02 by Xiaokang Chen, Chen, Xiaokang, Yuhui Yuan +5 · 36 citations
    Computer Science · #Advanced Image and Video Retrieval Techniques #Advanced Neural Network Applications #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Video Surveillance and Tracking Methods
  10. JanusFlow: Harmonizing Autoregression and Rectified Flow for Unified Multimodal Understanding and Generation
    2024/11/12 by Yiyang Ma, Xingchao Liu, Ma, Yiyang +24 · 39 citations
    Computer Science · #Natural Language Processing Techniques #Speech and dialogue systems #Topic Modeling
  11. 3D Sketch-aware Semantic Scene Completion via Semi-supervised Structure Prior
    2020/03/31 by Xiaokang Chen, Kwan-Yee Lin, Chen, Xiaokang +6 · 9 citations
    Computer Science · Engineering · #Advanced Vision and Imaging #Computer Graphics and Visualization Techniques #3D Shape Modeling and Analysis
  12. Real-time Neural Radiance Talking Portrait Synthesis via Audio-spatial Decomposition
    2022/11/22 by Jiaxiang Tang, Tang, Jiaxiang, Kaisiyuan Wang +15 · 10 citations
    Computer Science · Engineering · #Generative Adversarial Networks and Image Synthesis #Advanced Vision and Imaging #Human Motion and Animation
  13. Group DETR: Fast DETR Training with Group-Wise One-to-Many Assignment
    2022/07/26 by Qiang Chen, Chen, Qiang, Xiaokang Chen +9 · 5 citations
    Computer Science · Engineering · #Advanced Neural Network Applications #Domain Adaptation and Few-Shot Learning #Geophysical Methods and Applications
  14. Compressible-composable NeRF via Rank-residual Decomposition
    2022/05/30 by Jiaxiang Tang, Tang, Jiaxiang, Xiaokang Chen +5 · 4 citations
    Computer Science · #Advanced Vision and Imaging #Computer Graphics and Visualization Techniques #Advanced Image Processing Techniques
  15. Interactive Segment Anything NeRF with Feature Imitation
    2023/05/25 by Xiaokang Chen, Jiaxiang Tang, Chen, Xiaokang +7 · 3 citations
    Computer Science · Engineering · #Generative Adversarial Networks and Image Synthesis #3D Shape Modeling and Analysis #Computer Graphics and Visualization Techniques
  16. Tracing the source areas of detrital zircon and K-feldspar in the Yellow River Basin
    2024/02/21 by Xu Lin, Lin Xu, Qinmian Xu +9 · 3 citations
    Earth and Planetary Sciences · #Geological and Geochemical Analysis #Geology and Paleoclimatology Research #Geological formations and processes
  17. Point Scene Understanding via Disentangled Instance Mesh Reconstruction
    2022/03/31 by Jiaxiang Tang, Tang, Jiaxiang, Xiaokang Chen +5 · 1 citation
    Engineering · Earth and Planetary Sciences · #3D Shape Modeling and Analysis #3D Surveying and Cultural Heritage #Robotics and Sensor-Based Localization
  18. Improving Long Text Understanding with Knowledge Distilled from Summarization Model
    2024/05/08 by Yan Liu, Yazheng Yang, Liu, Yan +3 · 1 citation
    Computer Science · #Advanced Text Analysis Techniques #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Text and Document Classification Technologies #Topic Modeling