24 citations · 24 across the 2 of their papers we have counts for
2 papers
cs.CV2022★ 24 cited
EPIC-KITCHENS VISOR Benchmark: VIdeo Segmentations and Object Relations
Ahmad Darkhalil, Dandan Shan, Bin Zhu +6
We introduce VISOR, a new dataset of pixel annotations and a benchmark suite for segmenting hands and active objects in egocentric video. VISOR annotates videos from EPIC-KITCHENS,…
cs.CV2022
Hand-Object Interaction Reasoning
Jian Ma, Dima Damen
This paper proposes an interaction reasoning network for modelling spatio-temporal relationships between hands and objects in video. The proposed interaction unit utilises a Transf…