2 papers
cs.CV2025
Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG
Rakesh Raj Madavan, Akshat Kaimal, Hashim Faisal +1
An ensemble of trained multimodal encoders and vision-language models (VLMs) has become a standard approach for visual question answering (VQA) tasks. However, such models often fa…
cs.CV2024
GANESH: Generalizable NeRF for Lensless Imaging
Rakesh Raj Madavan, Akshat Kaimal, Badhrinarayanan K +4
Lensless imaging offers a significant opportunity to develop ultra-compact cameras by removing the conventional bulky lens system. However, without a focusing element, the sensor's…