vix.ing · top · new · best · stats · spec

Xiaoshan Yang

  1. Shifting More Attention to Visual Backbone: Query-modulated Refinement Networks for End-to-End Visual Grounding
    2022/03/29 by Jiabo Ye, Junfeng Tian, Ye, Jiabo +13 · 8 citations
    Computer Science · #Advanced Image and Video Retrieval Techniques #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Multimedia (cs.MM) #Multimodal Machine Learning Applications #Visual Attention and Saliency Detection
  2. Towards Visual Grounding: A Survey
    2024/12/28 by Linhui Xiao, Xiao, Linhui, Xiaoshan Yang +6 · 17 citations
    Computer Science · Social Sciences · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Geographic Information Systems Studies #Human Pose and Action Recognition #Multimodal Machine Learning Applications
  3. SgVA-CLIP: Semantic-guided Visual Adapting of Vision-Language Models for Few-shot Image Classification
    2022/11/28 by Peng, Fang, Xiaoshan Yang, Yang, Xiaoshan +3 · 7 citations
    Biochemistry, Genetics and Molecular Biology · Computer Science · #Cancer-related molecular mechanisms research #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Multimedia (cs.MM) #Multimodal Machine Learning Applications
  4. Multi-modal Queried Object Detection in the Wild
    2023/05/30 by Yifan Xu, Xu, Yifan, Mengdan Zhang +11 · 5 citations
    Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Text and Document Classification Technologies
  5. OneRef: Unified One-tower Expression Grounding and Segmentation with Mask Referring Modeling
    2024/10/10 by Linhui Xiao, Xiaoshan Yang, Xiao, Linhui +6 · 9 citations
    Computer Science · #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Speech and dialogue systems
  6. A Comprehensive Review of Few-shot Action Recognition
    2024/07/20 by Yuyang Wanyan, Wanyan, Yuyang, Xiaoshan Yang +5 · 6 citations
    Computer Science · #Advanced Neural Network Applications #Anomaly Detection Techniques and Applications #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Human Pose and Action Recognition
  7. Dynamic Hypergraph Convolutional Networks for Skeleton-Based Action Recognition
    2021/12/20 by Jinfeng Wei, Wei, Jinfeng, Yunxin Wang +9 · 2 citations
    Computer Science · Engineering · Medicine · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Human Pose and Action Recognition #Medical Imaging and Analysis #Stroke Rehabilitation and Recovery
  8. Look Before You Leap: A GUI-Critic-R1 Model for Pre-Operative Error Diagnosis in GUI Automation
    2025/06/05 by Yuyang Wanyan, Xi Zhang, Wanyan, Yuyang +21 · 10 citations
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Multimodal Machine Learning Applications #Speech and dialogue systems #Topic Modeling
  9. ECKPN: Explicit Class Knowledge Propagation Network for Transductive Few-shot Learning
    2021/06/16 by Chaofan Chen, Xiaoshan Yang, Chen, Chaofan +7 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning and ELM #Multimodal Machine Learning Applications
  10. Libra: Building Decoupled Vision System on Large Language Models
    2024/05/16 by Yifan Xu, Xiaoshan Yang, Xu, Yifan +5 · 2 citations
    Computer Science · Engineering · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Handwritten Text Recognition Techniques #Multimodal Machine Learning Applications #Robotics and Automated Systems
  11. VTM-Nav: Harnessing Cross-Episode Experience for Object-Goal Navigation with Hierarchical Visual-Topological Memory
    2026/07/16 by Xiaoran Xu, Yupeng Wu, Tianyu Xue +4
    #cs.CV #cs.AI