8 citations · 9 across the 3 of their papers we have counts for
3 papers
Beyond Logit Lens: Contextual Embeddings for Robust Hallucination Detection & Grounding in VLMs
Anirudh Phukan, Divyansh, Harshit Kumar Morj +3
The rapid development of Large Multimodal Models (LMMs) has significantly advanced multimodal understanding by harnessing the language abilities of Large Language Models (LLMs) and…
Towards Optimizing the Costs of LLM Usage
Shivanshu Shekhar, Tanishq Dubey, Koyel Mukherjee +3
Generative AI and LLMs in particular are heavily used nowadays for various document processing tasks such as question answering and summarization. However, different LLMs come with…
Social Media Ready Caption Generation for Brands
Himanshu Maheshwari, Koustava Goswami, Apoorv Saxena +1
Social media advertisements are key for brand marketing, aiming to attract consumers with captivating captions and pictures or logos. While previous research has focused on generat…