2 citations · 2 across the 3 of their papers we have counts for
3 papers
cs.CV2024
vid-TLDR: Training Free Token merging for Light-weight Video Transformer
Joonmyung Choi, Sanghyeok Lee, Jaewon Chu +2
Video Transformers have become the prevalent solution for various video downstream tasks with superior expressive power and flexibility. However, these video transformers suffer fr…
cs.CL2023
NuTrea: Neural Tree Search for Context-guided Multi-hop KGQA
Hyeong Kyu Choi, Seunghun Lee, Jaewon Chu +1
Multi-hop Knowledge Graph Question Answering (KGQA) is a task that involves retrieving nodes from a knowledge graph (KG) to answer natural language questions. Recent GNN-based appr…
cs.CV2023★ 2 cited
Open-vocabulary Video Question Answering: A New Benchmark for Evaluating the Generalizability of Video Question Answering Models
Dohwan Ko, Ji Soo Lee, Miso Choi +3
Video Question Answering (VideoQA) is a challenging task that entails complex multi-modal reasoning. In contrast to multiple-choice VideoQA which aims to predict the answer given s…