vix.ing · top · new · best · stats · spec

Dohwan Ko

  1. Large Language Models are Temporal and Causal Reasoners for Video Question Answering
    2023/10/24 by Dohwan Ko, Ko, Dohwan, Ji Soo Lee +7 · 10 citations
    Computer Science · #Multimodal Machine Learning Applications #Topic Modeling #Domain Adaptation and Few-Shot Learning
  2. MELTR: Meta Loss Transformer for Learning to Fine-tune Video Foundation Models
    2023/03/23 by Dohwan Ko, Joonmyung Choi, Ko, Dohwan +9 · 3 citations
    Computer Science · #Advanced Image and Video Retrieval Techniques #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Multimodal Machine Learning Applications