1 paper
Qiru Li, Ao Zhou, Zhiwei Jiang +4
Multi-label recognition with frozen Vision-Language Models (VLMs) is brittle under distribution shift: standard zero-shot inference scores labels independently, ignoring co-occurre…