collaborators
Showing cs.CVShow all

6 papers · 1 filter

cs.CV2026

Robust Onion: Peeling Open Vocab Object Detectors Under Noise

Priyank Pathak, Mukilan Karuppasamy, Aaditya Baranwal +2

The impact of real-world noise on Open Vocabulary Object Detectors (OV-ODs) remains poorly understood due to their architectural complexity. We present our comprehensive analysis R…

cs.CV2026

PHOEBI: An Open-World Benchmark for Bacterial Identification in Phase-Contrast Microscopy

Aaditya Baranwal, Md Jahid Hasan, Shruti Vyas

Optical microscopy enables rapid, label-free imaging of live bacteria and is the standard instrument for species identification across clinical, environmental, and industrial micro…

cs.CV2026

MolSight: Molecular Property Prediction with Images

Aaditya Baranwal, Akshaj Gupta, Yogesh S Rawat +1

Every molecule ever synthesised can be drawn as a 2D skeletal diagram, yet in modern property prediction this universally available representation has received less focus in favour…

cs.CV2026

BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs

Aaditya Baranwal, Vishal Yadav, Abhishek Rajora

While Vision-Language Models (VLMs) demonstrate remarkable zero-shot recognition capabilities across a diverse spectrum of multimodal tasks, it yet remains an open question whether…

cs.CV2025

Re:Verse -- Can Your VLM Read a Manga?

Aaditya Baranwal, Madhav Kataria, Naitik Agrawal +2

Current Vision Language Models (VLMs) demonstrate a critical gap between surface-level recognition and deep narrative reasoning when processing sequential visual storytelling. Thro…

cs.CV2025

SynSpill: Improved Industrial Spill Detection With Synthetic Data

Aaditya Baranwal, Abdul Mueez, Jason Voelker +2

Large-scale Vision-Language Models (VLMs) have transformed general-purpose visual recognition through strong zero-shot capabilities. However, their performance degrades significant…