2 citations · 2 across the 5 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2023
PixLore: A Dataset-driven Approach to Rich Image Captioning
Diego Bonilla-Salvador, Marcelino Martínez-Sober, Joan Vila-Francés +3
In the domain of vision-language integration, generating detailed image captions poses a significant challenge due to the lack of curated and rich datasets. This study introduces P…
cs.CV2023
Empirical study of the modulus as activation function in computer vision applications
Iván Vallés-Pérez, Emilio Soria-Olivas, Marcelino Martínez-Sober +3
In this work we propose a new non-monotonic activation function: the modulus. The majority of the reported research on nonlinearities is focused on monotonic functions. We empirica…