vix.ing · top · new · best · stats · spec

Li, Yunheng

  1. TeMO: Towards Text-Driven 3D Stylization for Multi-Object Meshes
    2023/12/07 by Zhang, Xuying, Yin, Bo-Wen, Chen, Yuming +4 · 3 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  2. SM3Det: A Unified Model for Multi-Modal Remote Sensing Object Detection
    2024/12/30 by Li, Yuxuan, Li, Xiang, Li, Yunheng +5 · 6 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Multimedia (cs.MM)
  3. A Decoupled Spatio-Temporal Framework for Skeleton-based Action Segmentation
    2023/12/10 by Li, Yunheng, Li, Zhongyu, Gao, Shanghua +3 · 2 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  4. High-Quality Mask Tuning Matters for Open-Vocabulary Segmentation
    2024/12/16 by Zeng, Quan-Sheng, Li, Yunheng, Zhou, Daquan +3 · 3 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  5. TempSamp-R1: Effective Temporal Sampling with Reinforcement Fine-Tuning for Video LLMs
    2025/09/22 by Yunheng Li, Li, Yunheng, Jing Cheng +10 · 6 citations
    Computer Science · Social Sciences · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Image and Video Quality Assessment #Multimedia Communication and Technology #Video Coding and Compression Technologies
  6. Unbiased Region-Language Alignment for Open-Vocabulary Dense Prediction
    2024/12/09 by Li, Yunheng, Li, Yuxuan, Zeng, Quansheng +3 · 3 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  7. Cascade-CLIP: Cascaded Vision-Language Embeddings Alignment for Zero-Shot Semantic Segmentation
    2024/06/02 by Yunheng Li, Li, Yunheng, Zhongyu Li +7 · 2 citations
    Computer Science · Medicine · #Multimodal Machine Learning Applications #Domain Adaptation and Few-Shot Learning #COVID-19 diagnosis using AI
  8. A Glimpse to Compress: Dynamic Visual Token Pruning for Large Vision-Language Models
    2025/08/03 by Zeng, Quan-Sheng, Li, Yunheng, Wang, Qilong +4 · 3 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences