14 citations · 14 across the 1 of their papers we have counts for
1 paper
Gouthaman KV, Athira Nambiar, Kancheti Sai Srinivas +1
Attention models are widely used in Vision-language (V-L) tasks to perform the visual-textual correlation. Humans perform such a correlation with a strong linguistic understanding…