19 citations
- Peking UniversityCN6 papers
- Hefei UniversityCN5 papers
- Zhejiang UniversityCN5 papers
- Longyan UniversityCN3 papers
- Sichuan UniversityCN3 papers
- The University of TokyoJP3 papers
- Beihang UniversityCN2 papers
- Chinese University of Hong Kong, ShenzhenCN2 papers
- Commonwealth Scientific and Industrial Research OrganisationAU2 papers
- Fujian University of TechnologyCN2 papers
- Kyoto UniversityJP2 papers
- Nanyang Technological UniversitySG2 papers
5 papers · 1 filter
MedP-CLIP: Medical CLIP with Region-Aware Prompt Integration
Jiahui Peng, He Yao, Jingwen Li +9
Contrastive Language-Image Pre-training (CLIP) has demonstrated outstanding performance in global image understanding and zero-shot transfer through large-scale text-image alignmen…
MSDS: Deep Structural Similarity with Multiscale Representation
Danling Kang, Xue-Hua Chen, Bin Liu +3
Deep-feature-based perceptual similarity models have demonstrated strong alignment with human visual perception in Image Quality Assessment (IQA). However, most existing approaches…
SGDFuse: SAM-Guided Diffusion Model for High-Fidelity Infrared and Visible Image Fusion
Xiaoyang Zhang, jinjiang Li, Guodong Fan +4
Infrared and visible image fusion (IVIF) is essential for integrating thermal saliency with textural details to support downstream perception. However, most existing approaches suf…
HAAP: Vision-context Hierarchical Attention Autoregressive with Adaptive Permutation for Scene Text Recognition
Honghui Chen, Yuhang Qiu, Jiabao Wang +2
Scene Text Recognition (STR) is challenging in extracting effective character representations from visual data when text is unreadable. Permutation language modeling (PLM) is intro…
Synthetic-to-Real Camouflaged Object Detection
Zhihao Luo, Luojun Lin, Zheng Lin
Due to the high cost of collection and labeling, there are relatively few datasets for camouflaged object detection (COD). In particular, for certain specialized categories, the av…