Dong, Shuwen
- VLM See, Robot Do: Human Demo Video to Robot Action Plan via Vision Language Model
2024/10/11 by Wang, Beichen, Zhang, Juexiao, Dong, Shuwen +2 · 19 citations
#Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Robotics (cs.RO)