vix.ing · top · new · best · stats · spec

Liu, Zhenyang

  1. ReasonGrounder: LVLM-Guided Hierarchical Feature Splatting for Open-Vocabulary 3D Visual Grounding and Reasoning
    2025/03/30 by Zhenyang Liu, Liu, Zhenyang, Yikai Wang +11 · 7 citations
    Computer Science · Engineering · #Advanced Neural Network Applications #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Robotics and Sensor-Based Localization
  2. TriVLA: A Triple-System-Based Unified Vision-Language-Action Model with Episodic World Modeling for General Robot Control
    2025/07/02 by Liu, Zhenyang, Gu, Yongchong, Zheng, Sixiao +3 · 3 citations
    #FOS: Computer and information sciences #Robotics (cs.RO)
  3. A Neural Representation Framework with LLM-Driven Spatial Reasoning for Open-Vocabulary 3D Visual Grounding
    2025/07/09 by Zhenyang Liu, Sixiao Zheng, Liu, Zhenyang +11 · 4 citations
    Computer Science · Engineering · #Computer Vision and Pattern Recognition (cs.CV) #Constraint Satisfaction and Optimization #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Robotics (cs.RO) #Spatial Cognition and Navigation