9 citations · 9 across the 4 of their papers we have counts for
6 papers
VisionTrap: Unanswerable Questions On Visual Data
Asir Saadat, Syem Aziz, Shahriar Mahmud +2
Visual Question Answering (VQA) has been a widely studied topic, with extensive research focusing on how VLMs respond to answerable questions based on real-world images. However, t…
Preemptive Hallucination Reduction: An Input-Level Approach for Multimodal Language Model
Nokimul Hasan Arif, Shadman Rabby, Md Hefzul Hossain Papon +1
Visual hallucinations in Large Language Models (LLMs), where the model generates responses that are inconsistent with the visual input, pose a significant challenge to their reliab…
DExNet: Combining Observations of Domain Adapted Critics for Leaf Disease Classification with Limited Data
Sabbir Ahmed, Md. Bakhtiar Hasan, Tasnim Ahmed +1
While deep learning-based architectures have been widely used for correctly detecting and classifying plant diseases, they require large-scale datasets to learn generalized feature…
Performance Analysis of Few-Shot Learning Approaches for Bangla Handwritten Character and Digit Recognition
Mehedi Ahamed, Radib Bin Kabir, Tawsif Tashwar Dipto +3
This study investigates the performance of few-shot learning (FSL) approaches in recognizing Bangla handwritten characters and numerals using limited labeled data. It demonstrates…
MangoLeafViT: Leveraging Lightweight Vision Transformer with Runtime Augmentation for Efficient Mango Leaf Disease Classification
Rafi Hassan Chowdhury, Sabbir Ahmed
Ensuring food safety is critical due to its profound impact on public health, economic stability, and global supply chains. Cultivation of Mango, a major agricultural product in se…
Performance Evaluation of Large Language Models in Bangla Consumer Health Query Summarization
Ajwad Abrar, Farzana Tabassum, Sabbir Ahmed
Consumer Health Queries (CHQs) in Bengali (Bangla), a low-resource language, often contain extraneous details, complicating efficient medical responses. This study investigates the…