2 papers
cs.CV2025
Leader360V: The Large-scale, Real-world 360 Video Dataset for Multi-task Learning in Diverse Environment
Weiming Zhang, Dingwen Xiao, Aobotao Dai +5
360 video captures the complete surrounding scenes with the ultra-large field of view of 360X180. This makes 360 scene understanding tasks, eg, segmentation and tracking, crucial f…
cs.CV2025
When Data Manipulation Meets Attack Goals: An In-depth Survey of Attacks for VLMs
Aobotao Dai, Xinyu Ma, Lei Chen +2
Vision-Language Models (VLMs) have gained considerable prominence in recent years due to their remarkable capability to effectively integrate and process both textual and visual in…