most citedSemantic Human Parsing via Scalable Semantic Transfer over Multiple Label Domains

1 citations · 3 across the 5 of their papers we have counts for

collaborators
Showing cs.CVShow all

5 papers · 1 filter

cs.CV2024

Narrative Action Evaluation with Prompt-Guided Multimodal Interaction

Shiyi Zhang, Sule Bai, Guangyi Chen +4

In this paper, we investigate a new problem called narrative action evaluation (NAE). NAE aims to generate professional commentary that evaluates the execution of an action. Unlike…

cs.CV2023

NPF-200: A Multi-Modal Eye Fixation Dataset and Method for Non-Photorealistic Videos

Ziyu Yang, Sucheng Ren, Zongwei Wu +4

Non-photorealistic videos are in demand with the wave of the metaverse, but lack of sufficient research studies. This work aims to take a step forward to understand how humans perc…

cs.CV2023

REC-MV: REconstructing 3D Dynamic Cloth from Monocular Videos

Lingteng Qiu, Guanying Chen, Jiapeng Zhou +3

Reconstructing dynamic 3D garment surfaces with open boundaries from monocular videos is an important problem as it provides a practical and low-cost solution for clothes digitizat…

cs.CV20231 cited

Semantic Human Parsing via Scalable Semantic Transfer over Multiple Label Domains

Jie Yang, Chaoqun Wang, Zhen Li +2

This paper presents Scalable Semantic Transfer (SST), a novel training paradigm, to explore how to leverage the mutual benefits of the data from different label domains (i.e. vario…

cs.CV20231 cited

NeMF: Inverse Volume Rendering with Neural Microflake Field

Youjia Zhang, Teng Xu, Junqing Yu +5

Recovering the physical attributes of an object's appearance from its images captured under an unknown illumination is challenging yet essential for photo-realistic rendering. Rece…