4 papers
Stage-wise Attention-Guided Region Sequencing for Adversarial Attacks on Large Vision-Language Models
Jaehyun Kwak, Nam Cao, Boryeong Cho +3
Targeted adversarial attacks on Large Vision-Language Models (LVLMs) test whether small image perturbations can steer model responses toward attacker-specified content. Under the s…
QuRe: Query-Relevant Retrieval through Hard Negative Sampling in Composed Image Retrieval
Jaehyun Kwak, Ramahdani Muhammad Izaaz Inhar, Se-Young Yun +1
Composed Image Retrieval (CIR) retrieves relevant images based on a reference image and accompanying text describing desired modifications. However, existing CIR methods only focus…
TV-LiVE: Training-Free, Text-Guided Video Editing via Layer Informed Vitality Exploitation
Min-Jung Kim, Dongjin Kim, Seokju Yun +1
Video editing has garnered increasing attention alongside the rapid progress of diffusion-based video generation models. As part of these advancements, there is a growing demand fo…
Synergistic Integration of Coordinate Network and Tensorial Feature for Improving Neural Radiance Fields from Sparse Inputs
Mingyu Kim, Jun-Seong Kim, Se-Young Yun +1
The multi-plane representation has been highlighted for its fast training and inference across static and dynamic neural radiance fields. This approach constructs relevant features…