1 citations · 1 across the 4 of their papers we have counts for
4 papers
VideoPure: Diffusion-based Adversarial Purification for Video Recognition
Kaixun Jiang, Zhaoyu Chen, Jiyuan Fu +3
Recent work indicates that video recognition models are vulnerable to adversarial examples, posing a serious security risk to downstream applications. However, current research has…
DeTrack: In-model Latent Denoising Learning for Visual Object Tracking
Xinyu Zhou, Jinglun Li, Lingyi Hong +4
Previous visual object tracking methods employ image-feature regression models or coordinate autoregression models for bounding box prediction. Image-feature regression methods hea…
X-Prompt: Multi-modal Visual Prompt for Video Object Segmentation
Pinxue Guo, Wanyun Li, Hao Huang +7
Multi-modal Video Object Segmentation (VOS), including RGB-Thermal, RGB-Depth, and RGB-Event, has garnered attention due to its capability to address challenging scenarios where tr…
LVOS: A Benchmark for Large-scale Long-term Video Object Segmentation
Lingyi Hong, Zhongying Liu, Wenchao Chen +9
Video object segmentation (VOS) aims to distinguish and track target objects in a video. Despite the excellent performance achieved by off-the-shell VOS models, existing VOS benchm…