activity
20162024
most citedPerformance Measures and a Data Set for Multi-Target, Multi-Camera Tracking

202 citations · 217 across the 32 of their papers we have counts for

collaborators
Showing cs.CVShow all

30 papers · 1 filter

cs.CV2024

BRIDGE: Bridging Gaps in Image Captioning Evaluation with Stronger Visual Cues

Sara Sarto, Marcella Cornia, Lorenzo Baraldi +1

Effectively aligning with human judgment when evaluating machine-generated image captions represents a complex yet intriguing challenge. Existing evaluation metrics like CIDEr or C…

cs.CV2024

Contrasting Deepfakes Diffusion via Contrastive Learning and Global-Local Similarities

Lorenzo Baraldi, Federico Cocchi, Marcella Cornia +2

Discerning between authentic content and that generated by advanced AI methods has become increasingly challenging. While previous research primarily addresses the detection of fak…

cs.CV20241 cited

Bringing Masked Autoencoders Explicit Contrastive Properties for Point Cloud Self-Supervised Learning

Bin Ren, Guofeng Mei, Danda Pani Paudel +6

Contrastive learning (CL) for Vision Transformers (ViTs) in image domains has achieved performance comparable to CL for traditional convolutional backbones. However, in 3D point cl…

cs.CV2024

Mask and Compress: Efficient Skeleton-based Action Recognition in Continual Learning

Matteo Mosconi, Andriy Sorokin, Aniello Panariello +6

The use of skeletal data allows deep learning models to perform action recognition efficiently and effectively. Herein, we believe that exploring this problem within the context of…

cs.CV2024

Binarizing Documents by Leveraging both Space and Frequency

Fabio Quattrini, Vittorio Pippi, Silvia Cascianelli +1

Document Image Binarization is a well-known problem in Document Analysis and Computer Vision, although it is far from being solved. One of the main challenges of this task is that…

cs.CV2024

AIGeN: An Adversarial Approach for Instruction Generation in VLN

Niyati Rawal, Roberto Bigazzi, Lorenzo Baraldi +1

In the last few years, the research interest in Vision-and-Language Navigation (VLN) has grown significantly. VLN is a challenging task that involves an agent following human instr…