32 citations · 99 across the 75 of their papers we have counts for
Showing cs.MMShow all
3 papers · 1 filter
cs.MM2023
Removing Interference and Recovering Content Imaginatively for Visible Watermark Removal
Yicheng Leng, Chaowei Fang, Gen Li +2
Visible watermarks, while instrumental in protecting image copyrights, frequently distort the underlying content, complicating tasks like scene interpretation and image editing. Vi…
cs.MM2023
MMAPS: End-to-End Multi-Grained Multi-Modal Attribute-Aware Product Summarization
Tao Chen, Ze Lin, Hui Li +4
Given the long textual product information and the product image, Multi-modal Product Summarization (MPS) aims to increase customers' desire to purchase by highlighting product cha…
cs.MM2021
Cross-Modal Self-Attention with Multi-Task Pre-Training for Medical Visual Question Answering
Haifan Gong, Guanqi Chen, Sishuo Liu +2
Due to the severe lack of labeled data, existing methods of medical visual question answering usually rely on transfer learning to obtain effective image feature representation and…