1 citations · 1 across the 2 of their papers we have counts for
1 paper · 1 filter
Pavana Pradeep, Krishna Kant, Suya Yu
Vision-Language Models (VLMs) offer the ability to generate high-level, interpretable descriptions of complex activities from images and videos, making them valuable for situationa…