2 citations · 2 across the 5 of their papers we have counts for
7 papers
ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation
Yuanjia Li, Tianyang Xu, Tao Zhou +3
Referring video object segmentation (RVOS) requires segmenting a target specified by natural language throughout a video. Recent agentic approaches combine multimodal large languag…
EvaNet: Towards More Efficient and Consistent Infrared and Visible Image Fusion Assessment
Chunyang Cheng, Tianyang Xu, Xiao-Jun Wu +4
Evaluation is essential in image fusion research, yet most existing metrics are directly borrowed from other vision tasks without proper adaptation. These traditional metrics, ofte…
UASTrack: A Unified Adaptive Selection Framework with Modality-Customization in Single Object Tracking
He Wang, Tianyang Xu, Zhangyong Tang +2
Multi-modal tracking is essential in single-object tracking (SOT), as different sensor types contribute unique capabilities to overcome challenges caused by variations in object ap…
Omni Survey for Multimodality Analysis in Visual Object Tracking
Zhangyong Tang, Tianyang Xu, Xuefeng Zhu +6
The development of smart cities has led to the generation of massive amounts of multi-modal data in the context of a range of tasks that enable a comprehensive monitoring of the sm…
Revisiting RGBT Tracking Benchmarks from the Perspective of Modality Validity: A New Benchmark, Problem, and Solution
Zhangyong Tang, Tianyang Xu, Zhenhua Feng +4
RGBT tracking draws increasing attention because its robustness in multi-modal warranting (MMW) scenarios, such as nighttime and adverse weather conditions, where relying on a sing…
Serial Over Parallel: Learning Continual Unification for Multi-Modal Visual Object Tracking and Benchmarking
Zhangyong Tang, Tianyang Xu, Xuefeng Zhu +4
Unifying multiple multi-modal visual object tracking (MMVOT) tasks draws increasing attention due to the complementary nature of different modalities in building robust tracking sy…