28 citations · 30 across the 3 of their papers we have counts for
8 papers
Inter-Homines: Distance-Based Risk Estimation for Human Safety
Matteo Fabbri, Fabio Lanzi, Riccardo Gasparini +3
In this document, we report our proposal for modeling the risk of possible contagiousity in a given area monitored by RGB cameras where people freely move and interact. Our system,…
A Novel Attention-based Aggregation Function to Combine Vision and Language
Matteo Stefanini, Marcella Cornia, Lorenzo Baraldi +1
The joint understanding of vision and language has been recently gaining a lot of attention in both the Computer Vision and Natural Language Processing communities, with the emerge…
Meshed-Memory Transformer for Image Captioning
Marcella Cornia, Matteo Stefanini, Lorenzo Baraldi +1
Transformer-based architectures represent the state of the art in sequence modeling tasks like machine translation and language understanding. Their applicability to multi-modal co…
SMArT: Training Shallow Memory-aware Transformers for Robotic Explainability
Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara
The ability to generate natural language explanations conditioned on the visual perception is a crucial step towards autonomous agents which can explain themselves and communicate…
Embodied Vision-and-Language Navigation with Dynamic Convolutional Filters
Federico Landi, Lorenzo Baraldi, Massimiliano Corsini +1
In Vision-and-Language Navigation (VLN), an embodied agent needs to reach a target destination with the only guidance of a natural language instruction. To explore the environment…
A Deep Learning based approach to VM behavior identification in cloud systems
Matteo Stefanini, Riccardo Lancellotti, Lorenzo Baraldi +1
Cloud computing data centers are growing in size and complexity to the point where monitoring and management of the infrastructure become a challenge due to scalability issues. A p…