29 citations · 36 across the 5 of their papers we have counts for
5 papers
Vision-and-Language Navigation via Causal Learning
Liuyi Wang, Zongtao He, Ronghao Dang +3
In the pursuit of robust and generalizable environment perception and language understanding, the ubiquitous challenge of dataset bias continues to plague vision-and-language navig…
Causality-based Cross-Modal Representation Learning for Vision-and-Language Navigation
Liuyi Wang, Zongtao He, Ronghao Dang +3
Vision-and-Language Navigation (VLN) has gained significant research interest in recent years due to its potential applications in real-world scenarios. However, existing VLN metho…
CLIPose: Category-Level Object Pose Estimation with Pre-trained Vision-Language Knowledge
Xiao Lin, Minghao Zhu, Ronghao Dang +5
Most of existing category-level object pose estimation methods devote to learning the object category information from point cloud modality. However, the scale of 3D datasets is li…
Multiple Thinking Achieving Meta-Ability Decoupling for Object Navigation
Ronghao Dang, Lu Chen, Liuyi Wang +3
We propose a meta-ability decoupling (MAD) paradigm, which brings together various object navigation methods in an architecture system, allowing them to mutually enhance each other…
Search for or Navigate to? Dual Adaptive Thinking for Object Navigation
Ronghao Dang, Liuyi Wang, Zongtao He +3
"Search for" or "Navigate to"? When finding an object, the two choices always come up in our subconscious mind. Before seeing the target, we search for the target based on experien…