4 papers · 1 filter
Tracking-by-detection in Multi-object Tracking: Survey and Experiments
Yujin Yang, Kyujin Shim, Kangwook Ko +1
Multi-object tracking (MOT) is an essential computer vision task that simultaneously tracks multiple objects in video sequences, with various applications in surveillance, autonomo…
AIM: Anchor Identity Features, Then Match for Multimodal Large Language Model Unlearning
Wonjun Lee, Jaehyuk Jang, Kangwook Ko +2
Multimodal large language models (MLLMs) can memorize identity-specific facts about people in their fine-tuning data, creating privacy risks when a person requests deletion. Existi…
T-VSS: Test-Time Visual Subspace Steering for Adversarial Robustness of Vision-Language Models
Jaehyuk Jang, Minseok Seo, Seungju Cho +2
Vision-language models (VLMs) achieve strong zero-shot recognition, but they remain highly vulnerable to adversarial perturbations. Recent test-time adaptations improve robustness…
VideoMamba: Spatio-Temporal Selective State Space Model
Jinyoung Park, Hee-Seon Kim, Kangwook Ko +2
We introduce VideoMamba, a novel adaptation of the pure Mamba architecture, specifically designed for video recognition. Unlike transformers that rely on self-attention mechanisms…