vix.ing · top · new · best · stats · spec

He, Mengwei

  1. B-VLLM: A Vision Large Language Model with Balanced Spatio-Temporal Tokens
    2024/12/13 by Lu, Zhuqiang, Yin, Zhenfei, He, Mengwei +4 · 2 citations
    #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences