3 citations · 4 across the 2 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2022★ 1 cited
Two-Stream Networks for Object Segmentation in Videos
Hannan Lu, Zhi Tian, Lirong Yang +2
Existing matching-based approaches perform video object segmentation (VOS) via retrieving support features from a pixel-level memory, while some pixels may suffer from lack of corr…
cs.CV2022★ 3 cited
Target-Driven Structured Transformer Planner for Vision-Language Navigation
Yusheng Zhao, Jinyu Chen, Chen Gao +5
Vision-language navigation is the task of directing an embodied agent to navigate in 3D scenes with natural language instructions. For the agent, inferring the long-term navigation…