110 citations · 176 across the 16 of their papers we have counts for
24 papers
Weak-shot Semantic Segmentation via Dual Similarity Transfer
Junjie Chen, Li Niu, Siyuan Zhou +3
Semantic segmentation is an important and prevalent task, but severely suffers from the high cost of pixel-level annotations when extending to more classes in wider applications. T…
Inharmonious Region Localization via Recurrent Self-Reasoning
Penghao Wu, Li Niu, Jing Liang +1
Synthetic images created by image editing operations are prevalent, but the color or illumination inconsistency between the manipulated region and background may make it unrealisti…
Inharmonious Region Localization with Auxiliary Style Feature
Penghao Wu, Li Niu, Liqing Zhang
With the prevalence of image editing techniques, users can create fantastic synthetic images, but the image quality may be compromised by the color/illumination discrepancy between…
From Representation to Reasoning: Towards both Evidence and Commonsense Reasoning for Video Question-Answering
Jiangtong Li, Li Niu, Liqing Zhang
Video understanding has achieved great success in representation learning, such as video caption, video object grounding, and video descriptive question-answer. However, current me…
Deep Video Harmonization with Color Mapping Consistency
Xinyuan Lu, Shengyuan Huang, Li Niu +2
Video harmonization aims to adjust the foreground of a composite video to make it compatible with the background. So far, video harmonization has only received limited attention an…
XYLayoutLM: Towards Layout-Aware Multimodal Networks For Visually-Rich Document Understanding
Zhangxuan Gu, Changhua Meng, Ke Wang +4
Recently, various multimodal networks for Visually-Rich Document Understanding(VRDU) have been proposed, showing the promotion of transformers by integrating visual and layout info…