Jongha Kim
M.S & Ph.D Integrated Student in MLV Lab, advised by Prof. Hyunwoo J. Kim.
Department of Computer Science and Engineering at Korea University, Seoul, Republic of Korea.
My research focuses on scaling what matters in multimodal large language model (MLLM) systems, moving beyond simply increasing data and compute. I pursue this goal across the full MLLM lifecycle through efficient model and agentic system design, high-quality supervision, post-training, and evidence-aware long-context inference.
I am currently a Research Scientist Intern at Adobe Research, Video AI Lab, working on language-guided video matting, mentored by Joon-Young Lee, Seoung Wug Oh, and David Seunghyun Yoon. I am also collaborating with Google Cloud AI, together with Jinsung Yoon.
For more information, please see my CV.
If you are interested in collaboration, opportunities, or just a quick chat, please feel free to reach out to me via email.
selected publications [full list]
(*) denotes equal contribution- WACVRelevance-aware Multi-context Contrastive Decoding for Retrieval-augmented Visual Question AnsweringIn IEEE/CVF Conference on Winter Conference on Applications of Computer Vision (WACV 2026)
- AAAITabFlash: Efficient Table Understanding with Progressive Question Conditioning and Token FocusingIn AAAI Conference on Artificial Intelligence (AAAI 2026)
- InfoScienceImproved Query Specialization for Transformer-based Visual Relationship DetectionIn Information Sciences (2026)