4 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.CV2021★ 4 cited
Label-Attention Transformer with Geometrically Coherent Objects for Image Captioning
Shikha Dubey, Farrukh Olimov, Muhammad Aasim Rafique +2
Automatic transcription of scene understanding in images and videos is a step towards artificial general intelligence. Image captioning is a nomenclature for describing meaningful…
cs.CV2021
Image Captioning using Multiple Transformers for Self-Attention Mechanism
Farrukh Olimov, Shikha Dubey, Labina Shrestha +2
Real-time image captioning, along with adequate precision, is the main challenge of this research field. The present work, Multiple Transformers for Self-Attention Mechanism (MTSM)…