3 papers
cs.CV2026
From Semantics to Pixels: Coarse-to-Fine Masked Autoencoders for Hierarchical Visual Understanding
Wenzhao Xiang, Yue Wu, Hongyang Yu +3
Self-supervised visual pre-training methods face an inherent tension: contrastive learning (CL) captures global semantics but loses fine-grained detail, while masked image modeling…
eess.IV2024
Dual-Stream Attention Network for Hyperspectral Image Unmixing
Yufang Wang, Wenmin Wu, Lin Qi +1
Hyperspectral image (HSI) contains abundant spatial and spectral information, making it highly valuable for unmixing. In this paper, we propose a Dual-Stream Attention Network (DSA…
cs.CV2023
Ranking-based Adaptive Query Generation for DETRs in Crowded Pedestrian Detection
Feng Gao, Jiaxu Leng, Ji Gan +1
DEtection TRansformer (DETR) and its variants (DETRs) have been successfully applied to crowded pedestrian detection, which achieved promising performance. However, we find that, i…