8 citations · 13 across the 4 of their papers we have counts for
2 papers
cs.CV2022
A Spatio-Temporal Attentive Network for Video-Based Crowd Counting
Marco Avvenuti, Marco Bongiovanni, Luca Ciampi +3
Automatic people counting from images has recently drawn attention for urban monitoring in modern Smart Cities due to the ubiquity of surveillance camera networks. Current computer…
cs.CV2022
ALADIN: Distilling Fine-grained Alignment Scores for Efficient Image-Text Matching and Retrieval
Nicola Messina, Matteo Stefanini, Marcella Cornia +4
Image-text matching is gaining a leading role among tasks involving the joint understanding of vision and language. In literature, this task is often used as a pre-training objecti…