13 citations · 13 across the 5 of their papers we have counts for
4 papers · 1 filter
Grafting Pre-trained Models for Multimodal Headline Generation
Lingfeng Qiao, Chen Wu, Ye Liu +3
Multimodal headline utilizes both video frames and transcripts to generate the natural language title of the videos. Due to a lack of large-scale, manually annotated data, the task…
Multi-Temporal Spatial-Spectral Comparison Network for Hyperspectral Anomalous Change Detection
Meiqi Hu, Chen Wu, Bo Du
Hyperspectral anomalous change detection has been a challenging task for its emphasis on the dynamics of small and rare objects against the prevalent changes. In this paper, we hav…
Weakly-Supervised Semantic Segmentation with Visual Words Learning and Hybrid Pooling
Lixiang Ru, Bo Du, Yibing Zhan +1
Weakly-Supervised Semantic Segmentation (WSSS) methods with image-level labels generally train a classification network to generate the Class Activation Maps (CAMs) as the initial…
Dual-Tasks Siamese Transformer Framework for Building Damage Assessment
Hongruixuan Chen, Edoardo Nemni, Sofia Vallecorsa +3
Accurate and fine-grained information about the extent of damage to buildings is essential for humanitarian relief and disaster response. However, as the most commonly used archite…