vix.ing · top · new · best · stats · spec

Dong, Shuwen

  1. VLM See, Robot Do: Human Demo Video to Robot Action Plan via Vision Language Model
    2024/10/11 by Wang, Beichen, Zhang, Juexiao, Dong, Shuwen +2 · 19 citations
    #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Robotics (cs.RO)