1 paper
Kyuho Lee, Euntae Kim, Jinwoo Choi +1
Video large language models (Video LLMs) have recently achieved strong performance on tasks such as captioning, summarization, and question answering. Many models and training meth…