1 citations · 1 across the 9 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
TDDN: Text-aligned Diffused DINO Network for Puzzle Understanding
Harsha Patnala, Debopriyo Banerjee, Ayush Sunil Munot +1
Structured visual reasoning, such as image puzzles, demands fine-grained visual perception, an ability current Vision Language Models (VLMs) lack. VLMs built on CLIP-based ViT back…
cs.CV2026
AD: Analysis and Detection of Adversarial Threats in Visual Perception for End-to-End Autonomous Driving Systems
Ishan Sahu, Somnath Hazra, Somak Aditya +1
End-to-end autonomous driving systems have achieved significant progress, yet their adversarial robustness remains largely underexplored. In this work, we conduct a closed-loop eva…