7 citations · 8 across the 2 of their papers we have counts for
2 papers
cs.CV2023★ 1 cited
Bridging the Gap: Exploring the Capabilities of Bridge-Architectures for Complex Visual Reasoning Tasks
Kousik Rajesh, Mrigank Raman, Mohammed Asad Karim +1
In recent times there has been a surge of multi-modal architectures based on Large Language Models, which leverage the zero shot generation capabilities of LLMs and project image e…
cs.CL2023★ 7 cited
HateProof: Are Hateful Meme Detection Systems really Robust?
Piush Aggarwal, Pranit Chawla, Mithun Das +4
Exploiting social media to spread hate has tremendously increased over the years. Lately, multi-modal hateful content such as memes has drawn relatively more traction than uni-moda…