3 papers
cs.CV2026
Boosting Segment Anything Model to Generalize Visually Non-Salient Scenarios
Guangqian Guo, Pengfei Chen, Yong Guo +3
Segment Anything Model (SAM), known for its remarkable zero-shot segmentation capabilities, has garnered significant attention in the community. Nevertheless, its performance is ch…
cs.CV2025
ReLayout: Integrating Relation Reasoning for Content-aware Layout Generation with Multi-modal Large Language Models
Jiaxu Tian, Xuehui Yu, Yaoxing Wang +3
Content-aware layout aims to arrange design elements appropriately on a given canvas to convey information effectively. Recently, the trend for this task has been to leverage large…
cs.CV2024
Why mamba is effective? Exploit Linear Transformer-Mamba Network for Multi-Modality Image Fusion
Chenguang Zhu, Shan Gao, Huafeng Chen +5
Multi-modality image fusion aims to integrate the merits of images from different sources and render high-quality fusion images. However, existing feature extraction and fusion met…