1 paper
Uday Bhaskar, Rishabh Bhattacharya, Avinash Patel +3
Foundation models, especially vision-language models (VLMs), offer compelling zero-shot object detection for applications like autonomous driving, a domain where manual labelling i…