1 paper · 1 filter
Sakib Ahammed, Xia Cui, Xinqi Fan +2
Modern vision models increasingly rely on rich semantic representations that extend beyond class labels to include descriptive concepts and contextual attributes. However, existing…