1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.CV2024
NAVERO: Unlocking Fine-Grained Semantics for Video-Language Compositionality
Chaofan Tao, Gukyeong Kwon, Varad Gunjal +7
We study the capability of Video-Language (VidL) models in understanding compositions between objects, attributes, actions and their relations. Composition understanding becomes pa…
cs.CL2023★ 1 cited
Generate then Select: Open-ended Visual Question Answering Guided by World Knowledge
Xingyu Fu, Sheng Zhang, Gukyeong Kwon +10
The open-ended Visual Question Answering (VQA) task requires AI models to jointly reason over visual and natural language inputs using world knowledge. Recently, pre-trained Langua…
eess.IV2022
Patient Aware Active Learning for Fine-Grained OCT Classification
Yash-yee Logan, Ryan Benkert, Ahmad Mustafa +2
This paper considers making active learning more sensible from a medical perspective. In practice, a disease manifests itself in different forms across patient cohorts. Existing fr…