2 citations · 4 across the 5 of their papers we have counts for
4 papers · 1 filter
Divide-and-Conquer: Tree-structured Strategy with Answer Distribution Estimator for Goal-Oriented Visual Dialogue
Shuo Cai, Xinzhe Han, Shuhui Wang
Goal-oriented visual dialogue involves multi-round interaction between artificial agents, which has been of remarkable attention due to its wide applications. Given a visual scene,…
Stable Attribute Group Editing for Reliable Few-shot Image Generation
Guanqi Ding, Xinzhe Han, Shuhui Wang +3
Few-shot image generation aims to generate data of an unseen category based on only a few samples. Apart from basic content generation, a bunch of downstream applications hopefully…
Multi-Attention Network for Compressed Video Referring Object Segmentation
Weidong Chen, Dexiang Hong, Yuankai Qi +5
Referring video object segmentation aims to segment the object referred by a given language expression. Existing works typically require compressed video bitstream to be decoded to…
Entity-enhanced Adaptive Reconstruction Network for Weakly Supervised Referring Expression Grounding
Xuejing Liu, Liang Li, Shuhui Wang +4
Weakly supervised Referring Expression Grounding (REG) aims to ground a particular target in an image described by a language expression while lacking the correspondence between ta…