2 papers
cs.CV2026
PatientVLM Meets DocVLM: Pre-Consultation Dialogue Between Vision-Language Models for Efficient Diagnosis
K Lokesh, Abhirama Subramanyam Penamakuri, Uday Agarwal +4
Traditionally, AI research in medical diagnosis has largely centered on image analysis. While this has led to notable advancements, the absence of patient-reported symptoms continu…
cs.CV2025
Aligning Moments in Time using Video Queries
Yogesh Kumar, Uday Agarwal, Manish Gupta +1
Video-to-video moment retrieval (Vid2VidMR) is the task of localizing unseen events or moments in a target video using a query video. This task poses several challenges, such as th…