Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
Video Summarization with Large Language Models
Min Jung Lee, Dayoung Gong, Minsu Cho
The exponential increase in video content poses significant challenges in terms of efficient navigation, search, and retrieval, thus requiring advanced video summarization techniqu…
cs.CV2025
Locality-Aware Zero-Shot Human-Object Interaction Detection
Sanghyun Kim, Deunsol Jung, Minsu Cho
Recent methods for zero-shot Human-Object Interaction (HOI) detection typically leverage the generalization ability of large Vision-Language Model (VLM), i.e., CLIP, on unseen cate…
cs.CV2024
Burst Image Super-Resolution with Base Frame Selection
Sanghyun Kim, Min Jung Lee, Woohyeok Kim +4
Burst image super-resolution has been a topic of active research in recent years due to its ability to obtain a high-resolution image by using complementary information between mul…