11 papers
HiSC: Hierarchical Spatial Clustering Token Compression for Efficient 3D Scene Understanding
Jiuhe Qu, Yingping Liang, Ying Fu
3D vision-language models (3D VLMs) enable spatial reasoning over multi-view scenes but suffer from substantial token redundancy due to duplicated observations and large uninformat…
Learning Semantic-Robust Change Detection via Semantic-Invariant Self-Distillation
Jiuhe Qu, Yingping Liang, Ying Fu
Change detection aims to identify semantic changes between remote sensing images. However, features from models are easily disturbed by non-semantic variations, such as illuminatio…
AquaStereo: Enabling Underwater Stereo Matching via Depth-Conditioned Diffusion and Geometry Self-Distillation
Qizhe Wei, Yingping Liang, Shaodi You +1
Learning-based stereo matching models struggle in underwater environments due to scarce in-domain data and the difficulty of extracting discriminative correspondences from degraded…
Boosting Text-Driven Video Segmentation via Geometry-Aware Distillation
Tianyu Zhu, Yingping Liang, Hesong Li +1
Text-driven Referring Video Object Segmentation (RVOS) aims to locate and segment target objects in videos given natural language. However, existing models are typically trained on…
A large-scale nanocrystal database with aligned synthesis and properties enabling generative inverse design
Kai Gu, Yingping Liang, Senliang Peng +3
The synthesis of nanocrystals has been highly dependent on trial-and-error, due to the complex correlation between synthesis parameters and physicochemical properties. Although dee…
Learning Dense Feature Matching via Lifting Single 2D Image to 3D Space
Yingping Liang, Yutao Hu, Wenqi Shao +1
Feature matching plays a fundamental role in many computer vision tasks, yet existing methods heavily rely on scarce and clean multi-view image collections, which constrains their…