activity
20182022
most citedM-VAD Names: a Dataset for Video Captioning with Naming

28 citations · 30 across the 5 of their papers we have counts for

collaborators
Showing cs.CVShow all

5 papers · 1 filter

cs.CV2021

Multi-Category Mesh Reconstruction From Image Collections

Alessandro Simoni, Stefano Pini, Roberto Vezzani +1

Recently, learning frameworks have shown the capability of inferring the accurate shape, pose, and texture of an object from a single RGB image. However, current methods are traine…

cs.CV20212 cited

SHREC 2021: Track on Skeleton-based Hand Gesture Recognition in the Wild

Ariel Caputo, Andrea Giachetti, Simone Soso +16

Gesture recognition is a fundamental tool to enable novel interaction paradigms in a variety of application scenarios like Mixed Reality environments, touchless public kiosks, ente…

cs.CV201928 cited

M-VAD Names: a Dataset for Video Captioning with Naming

Stefano Pini, Marcella Cornia, Federico Bolelli +2

Current movie captioning architectures are not capable of mentioning characters with their proper name, replacing them with a generic "someone" tag. The lack of movie description d…

cs.CV2018

Learn to See by Events: Color Frame Synthesis from Event and RGB Cameras

Stefano Pini, Guido Borghi, Roberto Vezzani

Event cameras are biologically-inspired sensors that gather the temporal evolution of the scene. They capture pixel-wise brightness variations and output a corresponding stream of…

cs.CV2018

Learning to Generate Facial Depth Maps

Stefano Pini, Filippo Grazioli, Guido Borghi +2

In this paper, an adversarial architecture for facial depth map estimation from monocular intensity images is presented. By following an image-to-image approach, we combine the adv…