24 citations · 25 across the 2 of their papers we have counts for
2 papers
cs.CV2022★ 24 cited
EPIC-KITCHENS VISOR Benchmark: VIdeo Segmentations and Object Relations
Ahmad Darkhalil, Dandan Shan, Bin Zhu +6
We introduce VISOR, a new dataset of pixel annotations and a benchmark suite for segmenting hands and active objects in egocentric video. VISOR annotates videos from EPIC-KITCHENS,…
cs.CV2022★ 1 cited
When Did It Happen? Duration-informed Temporal Localization of Narrated Actions in Vlogs
Oana Ignat, Santiago Castro, Yuhang Zhou +3
We consider the task of temporal human action localization in lifestyle vlogs. We introduce a novel dataset consisting of manual annotations of temporal localization for 13,000 nar…