1 paper
Diego Bonilla-Salvador, Marcelino Martínez-Sober, Joan Vila-Francés +3
In the domain of vision-language integration, generating detailed image captions poses a significant challenge due to the lack of curated and rich datasets. This study introduces P…