5 papers
Multi-camera Torso Pose Estimation using Graph Neural Networks
Daniel Rodriguez-Criado, Pilar Bachiller, Pablo Bustos +2
Estimating the location and orientation of humans is an essential skill for service and assistive robots. To achieve a reliable estimation in a wide area such as an apartment, mult…
Look and Listen: A Multi-modality Late Fusion Approach to Scene Classification for Autonomous Machines
Jordan J. Bird, Diego R. Faria, Cristiano Premebida +2
The novelty of this study consists in a multi-modality approach to scene classification, where image and audio complement each other in a process of deep late fusion. The approach…
Domain Adaptation for Reinforcement Learning on the Atari
Thomas Carr, Maria Chli, George Vogiatzis
Deep reinforcement learning agents have recently been successful across a variety of discrete and continuous control tasks; however, they can be slow to train and require a large n…
How to Read Paintings: Semantic Art Understanding with Multi-Modal Retrieval
Noa Garcia, George Vogiatzis
Automatic art analysis has been mostly focused on classifying artworks into different artistic styles. However, understanding an artistic representation involves more complex proce…
Dress like a Star: Retrieving Fashion Products from Videos
Noa Garcia, George Vogiatzis
This work proposes a system for retrieving clothing and fashion products from video content. Although films and television are the perfect showcase for fashion brands to promote th…