vix.ing · top · new · best · stats · spec

Hongjun Zhang

  1. DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning
    DeepSeek-R1 shows an LLM can learn strong step-by-step reasoning from pure reinforcement learning, with no human-labeled reasoning examples.
    2025/01/22 by DeepSeek-AI, Daya Guo, Guo, Daya +404 · 93 voices · 2776 citations
    Computer Science · #Reinforcement Learning in Robotics #Data Stream Mining Techniques #Explainable Artificial Intelligence (XAI)
  2. An attempt to generate new bridge types from latent space of denoising diffusion Implicit model
    2024/02/11 by Hongjun Zhang, Zhang, Hongjun · 1 citation
    Computer Science · Engineering · #Artificial Intelligence (cs.AI) #Blasting Impact and Analysis #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Image Processing and 3D Reconstruction #Machine Learning (cs.LG) #Neural Networks and Applications