3 papers
cs.CV2026
Learning Question-Aware Keyframe Selection with Synthetic Supervision for Video Question Answering
Minchan Kwon, Hyounguk Shon, Junmo Kim
Large multimodal models (LMMs) have recently demonstrated remarkable performance in video question answering (VideoQA), yet reasoning over video remains challenging due to high inf…
cs.CV2025
DMQ: Dissecting Outliers of Diffusion Models for Post-Training Quantization
Dongyeun Lee, Jiwan Hur, Hyounguk Shon +2
Diffusion models have achieved remarkable success in image generation but come with significant computational costs, posing challenges for deployment in resource-constrained enviro…
cs.CV2025
SFLD: Reducing the content bias for AI-generated Image Detection
Seoyeon Gye, Junwon Ko, Hyounguk Shon +2
Identifying AI-generated content is critical for the safe and ethical use of generative AI. Recent research has focused on developing detectors that generalize to unknown generator…