vix.ing · top · new · best · stats · spec

Shangguan, Ziyao

  1. MMVU: Measuring Expert-Level Multi-Discipline Video Understanding
    2025/01/21 by Yilun Zhao, Zhao, Yilun, Haowei Zhang +34 · 43 citations
    Medicine · #Artificial Intelligence (cs.AI) #Clinical Reasoning and Diagnostic Skills #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Radiology practices and education
  2. TOMATO: Assessing Visual Temporal Reasoning Capabilities in Multimodal Foundation Models
    2024/10/30 by Ziyao Shangguan, Chuhan Li, Shangguan, Ziyao +11 · 1 voice · 17 citations
    Computer Science · Social Sciences · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Geographic Information Systems Studies #Semantic Web and Ontologies #Speech and dialogue systems #cs.AI #cs.CL #cs.CV
  3. M3SciQA: A Multi-Modal Multi-Document Scientific QA Benchmark for Evaluating Foundation Models
    2024/11/06 by Chuhan Li, Li, Chuhan, Ziyao Shangguan +9 · 9 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Educational Technology and Assessment #FOS: Computer and information sciences #Semantic Web and Ontologies
  4. INSIGHT: INference-time Sequence Introspection for Generating Help Triggers in Vision-Language-Action Models
    2025/10/01 by Ulas Berk Karli, Karli, Ulas Berk, Ziyao Shangguan +3 · 1 citation
    Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Robotics (cs.RO) #Speech and dialogue systems