vix.ing · top · new · best · stats · spec

Riddell, James

  1. LEO: Boosting Mixture of Vision Encoders for Multimodal Large Language Models
    2025/01/13 by Azadani, Mozhgan Nasr, Riddell, James, Sedwards, Sean +1 · 4 citations
    #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences