4 citations · 4 across the 4 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2024
TINA: Think, Interaction, and Action Framework for Zero-Shot Vision Language Navigation
Dingbang Li, Wenzhou Chen, Xin Lin
Zero-shot navigation is a critical challenge in Vision-Language Navigation (VLN) tasks, where the ability to adapt to unfamiliar instructions and to act in unknown environments is…
cs.CV2024
TD^2-Net: Toward Denoising and Debiasing for Dynamic Scene Graph Generation
Xin Lin, Chong Shi, Yibing Zhan +3
Dynamic scene graph generation (SGG) focuses on detecting objects in a video and determining their pairwise relationships. Existing dynamic SGG methods usually suffer from several…