2 papers
cs.CV2025
DiVE-k: Differential Visual Reasoning for Fine-grained Image Recognition
Raja Kumar, Arka Sadhu, Ram Nevatia
Large Vision Language Models (LVLMs) possess extensive text knowledge but struggles to utilize this knowledge for fine-grained image recognition, often failing to differentiate bet…
cs.CV2023
Efficient Feature Distillation for Zero-shot Annotation Object Detection
Zhuoming Liu, Xuefeng Hu, Ram Nevatia
We propose a new setting for detecting unseen objects called Zero-shot Annotation object Detection (ZAD). It expands the zero-shot object detection setting by allowing the novel ob…