7 citations · 7 across the 3 of their papers we have counts for
5 papers · 1 filter
MetaDance: Few-shot Dancing Video Retargeting via Temporal-aware Meta-learning
Yuying Ge, Yibing Song, Ruimao Zhang +1
Dancing video retargeting aims to synthesize a video that transfers the dance movements from a source video to a target person. Previous work need collect a several-minute-long vid…
MetaCloth: Learning Unseen Tasks of Dense Fashion Landmark Detection from a Few Samples
Yuying Ge, Ruimao Zhang, Ping Luo
Recent advanced methods for fashion landmark detection are mainly driven by training convolutional neural networks on large-scale fashion datasets, which has a large number of anno…
End-to-End Dense Video Captioning with Parallel Decoding
Teng Wang, Ruimao Zhang, Zhichao Lu +3
Dense video captioning aims to generate multiple associated captions with their temporal locations from the video. Previous methods follow a sophisticated "localize-then-describe"…
PolarMask++: Enhanced Polar Representation for Single-Shot Instance Segmentation and Beyond
Enze Xie, Wenhai Wang, Mingyu Ding +2
Reducing the complexity of the pipeline of instance segmentation is crucial for real-world applications. This work addresses this issue by introducing an anchor-box free and single…
Polygon-free: Unconstrained Scene Text Detection with Box Annotations
Weijia Wu, Enze Xie, Ruimao Zhang +3
Although a polygon is a more accurate representation than an upright bounding box for text detection, the annotations of polygons are extremely expensive and challenging. Unlike ex…