jongha-crop.jpeg

Jongha Kim

M.S & Ph.D Integrated Student in MLV Lab, advised by Prof. Hyunwoo J. Kim.

Department of Computer Science and Engineering at Korea University, Seoul, Republic of Korea.

My research focuses on scaling what matters in multimodal large language model (MLLM) systems, moving beyond simply increasing data and compute. I pursue this goal across the full MLLM lifecycle through efficient model and agentic system design, high-quality supervision, post-training, and evidence-aware long-context inference.

I am currently a Research Scientist Intern at Adobe Research, Video AI Lab, working on language-guided video matting, mentored by Joon-Young Lee, Seoung Wug Oh, and David Seunghyun Yoon. I am also collaborating with Google Cloud AI, together with Jinsung Yoon.

For more information, please see my CV.

If you are interested in collaboration, opportunities, or just a quick chat, please feel free to reach out to me via email.

selected publications [full list]

(*) denotes equal contribution

  1. WACV
    Relevance-aware Multi-context Contrastive Decoding for Retrieval-augmented Visual Question Answering
    Jongha Kim, Byungoh Ko, Jeehye Na, Jinsung Yoon, and Hyunwoo J Kim
    In IEEE/CVF Conference on Winter Conference on Applications of Computer Vision (WACV 2026)
  2. AAAI
    TabFlash: Efficient Table Understanding with Progressive Question Conditioning and Token Focusing
    Jongha Kim, Minseong Bae, Sanghyeok Lee, Jinsung Yoon, and Hyunwoo J Kim
    In AAAI Conference on Artificial Intelligence (AAAI 2026)
  3. InfoScience
    Improved Query Specialization for Transformer-based Visual Relationship Detection
    Jongha Kim, Jihwan Park, Jinyoung Park, Jinyoung Kim, Sehyung Kim, and Hyunwoo J Kim
    In Information Sciences (2026)
  4. AAAI
    VidChain: Chain-of-Tasks with Metric-based Direct Preference Optimization for Dense Video Captioning
    Ji Soo Lee*, Jongha Kim*, Jeehye Na, Jinyoung Park, and Hyunwoo J Kim
    In AAAI Conference on Artificial Intelligence (AAAI 2025)
  5. CVPR
    Groupwise Query Specialization and Quality-Aware Multi-Assignment for Transformer-based Visual Relationship Detection
    Jongha Kim*, Jihwan Park*, Jinyoung Park*, Jinyoung Kim, Sehyung Kim, and Hyunwoo J Kim
    In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR 2024)