4 papers
Physics-Based Benchmarking Metrics for Multimodal Synthetic Images
Kishor Datta Gupta, Marufa Kamal, Md. Mahfuzur Rahman +3
Current state of the art measures like BLEU, CIDEr, VQA score, SigLIP-2 and CLIPScore are often unable to capture semantic or structural accuracy, especially for domain-specific or…
VLCE: A Knowledge-Enhanced Framework for Image Description in Disaster Assessment
Md. Mahfuzur Rahman, Kishor Datta Gupta, Marufa Kamal +5
General-purpose vision-language models (VLMs) such as LLaVA and QwenVL produce descriptions of disaster imagery that lack domain-specific vocabulary and actionable detail. We propo…
Advanced Tool Learning and Selection System (ATLASS): A Closed-Loop Framework Using LLM
Mohd Ariful Haque, Justin Williams, Sunzida Siddique +4
The combination of LLM agents with external tools enables models to solve complex tasks beyond their knowledge base. Human-designed tools are inflexible and restricted to solutions…
SOK: Exploring Hallucinations and Security Risks in AI-Assisted Software Development with Insights for LLM Deployment
Ariful Haque, Sunzida Siddique, Md. Mahfuzur Rahman +5
The integration of Large Language Models (LLMs) such as GitHub Copilot, ChatGPT, Cursor AI, and Codeium AI into software development has revolutionized the coding landscape, offeri…