Showing cs.CVShow all
2 papers · 1 filter
cs.CV2024
Multimodal video analysis for crowd anomaly detection using open access tourism cameras
Alejandro Dionis-Ros, Joan Vila-Francés, Rafael Magdalena-Benedicto +2
In this article, we propose the detection of crowd anomalies through the extraction of information in the form of time series from video format using a multimodal approach. Through…
cs.CV2023
PixLore: A Dataset-driven Approach to Rich Image Captioning
Diego Bonilla-Salvador, Marcelino Martínez-Sober, Joan Vila-Francés +3
In the domain of vision-language integration, generating detailed image captions poses a significant challenge due to the lack of curated and rich datasets. This study introduces P…