vix.ing · top · new · best · stats · spec

Junhyeok Kim

  1. Interpreting Attention Heads for Image-to-Text Information Flow in Large Vision-Language Models
    2025/09/22 by Jin-Yeong Kim, Kim, Jinyeong, Seil Kang +7 · 2 citations
    Computer Science · #Multimodal Machine Learning Applications
  2. v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning
    2025/05/24 by Jiwan Chung, Junhyeok Kim, Chung, Jiwan +8 · 3 citations
    Computer Science · #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Speech and dialogue systems