I am a master's student in the MLV (Machine Learning & Vision) Lab at Korea University, South Korea, advised by Prof. Hyunwoo J. Kim. I received my B.S. in Computer Science and Engineering from Korea University. My research interests lie in multi-modal understanding, especially in video understanding with foundation models.
Most recent publications on Google Scholar.
* indicates equal contribution.
Captioning for Text-Video Retrieval via DualGroup-Direct Preference Optimization
Ji Soo Lee, Byungoh Ko, Jaewon Cho, Howoong Lee, Jaewoon Byun, Hyunwoo J. Kim
EMNLP 2025 findings
Bidirectional Likelihood Estimation with Multi-Modal Large Language Models for Text-Video Retrieval
Dohwan Ko*, Ji Soo Lee*, Minhyuk Choi, Zihang Meng, Hyunwoo J. Kim
ICCV 2025 Highlight
VidChain: Chain-of-Tasks with Metric-based Direct Preference Optimization for Dense Video Captioning
Ji Soo Lee*, Jongha Kim*, Jeehye Na, Jinyoung Park, Hyunwoo J. Kim
AAAI 2025
Large Language Models are Temporal and Causal Reasoners for Video Question Answering
Dohwan Ko*, Ji Soo Lee*, Wooyoung Kang, Byungseok Roh, Hyunwoo J. Kim
EMNLP 2023
Open-vocabulary Video Question Answering: A New Benchmark for Evaluating the Generalizability of Video Question Answering Models
Dohwan Ko, Ji Soo Lee, Miso Choi, Jaewon Chu, Jihwan Park, Hyunwoo J. Kim
ICCV 2023
Full Resume in PDF.