most citedRGB-D Video Object Segmentation via Enhanced Multi-store Feature Memory

4 citations · 6 across the 4 of their papers we have counts for

collaborators

5 papers

cs.CV2025

SwiTrack: Tri-State Switch for Cross-Modal Object Tracking

Boyue Xu, Ruichao Hou, Tongwei Ren +3

Cross-modal object tracking (CMOT) is an emerging task that maintains target consistency while the video stream switches between different modalities, with only one modality availa…

cs.CV2025

MTNet: Learning modality-aware representation with transformer for RGBT tracking

Ruichao Hou, Boyue Xu, Tongwei Ren +1

The ability to learn robust multi-modality representation has played a critical role in the development of RGBT tracking. However, the regular fusion paradigm and the invariable tr…

cs.CV2025

Learning Frequency and Memory-Aware Prompts for Multi-Modal Object Tracking

Boyue Xu, Ruichao Hou, Tongwei Ren +3

Prompt-learning-based multi-modal trackers have made strong progress by using lightweight visual adapters to inject auxiliary-modality cues into frozen foundation models. However,…

cs.CV20252 cited

RGB-D Tracking via Hierarchical Modality Aggregation and Distribution Network

Boyue Xu, Yi Xu, Ruichao Hou +3

The integration of dual-modal features has been pivotal in advancing RGB-Depth (RGB-D) tracking. However, current trackers are less efficient and focus solely on single-level featu…

cs.CV20254 cited

RGB-D Video Object Segmentation via Enhanced Multi-store Feature Memory

Boyue Xu, Ruichao Hou, Tongwei Ren +1

The RGB-Depth (RGB-D) Video Object Segmentation (VOS) aims to integrate the fine-grained texture information of RGB with the spatial geometric clues of depth modality, boosting the…