50 citations · 96 across the 8 of their papers we have counts for
1 paper · 1 filter
Anindya Mondal, Sauradip Nag, Anjan Dutta
We present ABACUS, a unified vision-language model that jointly addresses object counting, crowd counting, referring-expression counting, and count-faithful image generation within…