2 citations · 2 across the 20 of their papers we have counts for
4 papers · 1 filter
A Data Efficiency Study of Synthetic Fog for Object Detection Using the Clear2Fog Pipeline
Mohamed Ahmed Mohamed, Xiaowei Huang
Object detection in adverse weather is critical for the safety of autonomous vehicles; however, the scarcity of labelled, real-world foggy data remains a significant bottleneck. In…
Spatial-DISE: A Unified Benchmark for Evaluating Spatial Reasoning in Vision-Language Models
Xinmiao Huang, Qisong He, Zhenglin Huang +5
Spatial reasoning ability is crucial for Vision Language Models (VLMs) to support real-world applications in diverse domains including robotics, augmented reality, and autonomous n…
TAIJI: Textual Anchoring for Immunizing Jailbreak Images in Vision Language Models
Xiangyu Yin, Yi Qi, Jinwei Hu +5
Vision Language Models (VLMs) have demonstrated impressive inference capabilities, but remain vulnerable to jailbreak attacks that can induce harmful or unethical responses. Existi…
CeTAD: Towards Certified Toxicity-Aware Distance in Vision Language Models
Xiangyu Yin, Jiaxu Liu, Zhen Chen +4
Recent advances in large vision-language models (VLMs) have demonstrated remarkable success across a wide range of visual understanding tasks. However, the robustness of these mode…