4 citations · 5 across the 3 of their papers we have counts for
3 papers
cs.CV2024★ 1 cited
Aligning Knowledge Graph with Visual Perception for Object-goal Navigation
Nuo Xu, Wen Wang, Rong Yang +6
Object-goal navigation is a challenging task that requires guiding an agent to specific objects based on first-person visual observations. The ability of agent to comprehend its su…
cs.SD2022★ 4 cited
End-To-End Audiovisual Feature Fusion for Active Speaker Detection
Fiseha B. Tesema, Zheyuan Lin, Shiqiang Zhu +3
Active speaker detection plays a vital role in human-machine interaction. Recently, a few end-to-end audiovisual frameworks emerged. However, these models' inference time was not e…
cs.CV2022
TGRMPT: A Head-Shoulder Aided Multi-Person Tracker and a New Large-Scale Dataset for Tour-Guide Robot
Wen Wang, Shunda Hu, Shiqiang Zhu +5
A service robot serving safely and politely needs to track the surrounding people robustly, especially for Tour-Guide Robot (TGR). However, existing multi-object tracking (MOT) or…