vix.ing · top · new · best · stats · spec

Chen, Xuweiyi

  1. LLM-Grounder: Open-Vocabulary 3D Visual Grounding with Large Language Model as an Agent
    2023/09/21 by Yang, Jianing, Chen, Xuweiyi, Qian, Shengyi +4 · 19 citations
    #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Robotics (cs.RO)
  2. Multi-Object Hallucination in Vision-Language Models
    2024/07/08 by Xuweiyi Chen, Ziqiao Ma, Chen, Xuweiyi +13 · 18 citations
    Biochemistry, Genetics and Molecular Biology · Medicine · Psychology · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #Epilepsy research and treatment #FOS: Computer and information sciences #Pharmacological Receptor Mechanisms and Effects #Psychedelics and Drug Studies
  3. 3D-GRAND: A Million-Scale Dataset for 3D-LLMs with Better Grounding and Less Hallucination
    2024/06/07 by Jianing Yang, Yang, Jianing, Xuweiyi Chen +11 · 10 citations
    Medicine · Engineering · #Medical Imaging Techniques and Applications #Advanced X-ray and CT Imaging #Advanced Surface Polishing Techniques
  4. Open Vocabulary Monocular 3D Object Detection
    2024/11/25 by Yao, Jin, Gu, Hao, Chen, Xuweiyi +2 · 4 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  5. UniCtrl: Improving the Spatiotemporal Consistency of Text-to-Video Diffusion Models via Training-Free Unified Attention Control
    2024/03/04 by Tian Xia, Xuweiyi Chen, Xia, Tian +3 · 2 citations
    Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Multimodal Machine Learning Applications
  6. Frame In-N-Out: Unbounded Controllable Image-to-Video Generation
    2025/05/27 by Boyang Wang, Xuweiyi Chen, Wang, Boyang +5 · 4 citations
    Computer Science · Engineering · #Advanced Vision and Imaging #Image Processing Techniques and Applications #Computer Graphics and Visualization Techniques
  7. Probing the Mid-level Vision Capabilities of Self-Supervised Learning
    2024/11/25 by Xuweiyi Chen, Markus Marks, Chen, Xuweiyi +3 · 2 citations
    Social Sciences · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Online and Blended Learning
  8. SAB3R: Semantic-Augmented Backbone in 3D Reconstruction
    2025/06/02 by Xuweiyi Chen, Tian Xia, Chen, Xuweiyi +9 · 2 citations
    Computer Science · Engineering · #3D Shape Modeling and Analysis #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Robotics and Sensor-Based Localization
  9. 4D-LRM: Large Space-Time Reconstruction Model From and To Any View at Any Time
    2025/06/23 by Ma, Ziqiao, Chen, Xuweiyi, Yu, Shoubin +10 · 2 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences