activity
20232026
most citedMVImgNet2.0: A Larger-scale Dataset of Multi-view Images

1 citations · 1 across the 3 of their papers we have counts for

collaborators

6 papers

cs.CV2026

ReImagine: Rethinking Controllable High-Quality Human Video Generation via Image-First Synthesis

Zhengwentai Sun, Keru Zheng, Chenghong Li +7

Human video generation remains challenging due to the difficulty of jointly modeling human appearance, motion, and camera viewpoint under limited multi-view data. Existing methods…

cs.CV2025

MV-Performer: Taming Video Diffusion Model for Faithful and Synchronized Multi-view Performer Synthesis

Yihao Zhi, Chenghong Li, Hongjie Liao +6

Recent breakthroughs in video generation, powered by large-scale datasets and diffusion techniques, have shown that video diffusion models can function as implicit 4D novel view sy…

cs.CV2025

MVHumanNet++: A Large-scale Dataset of Multi-view Daily Dressing Human Captures with Richer Annotations for 3D Human Digitization

Chenghong Li, Hongjie Liao, Yihao Zhi +5

In this era, the success of large language models and text-to-image models can be attributed to the driving force of large-scale datasets. However, in the realm of 3D vision, while…

cs.CV2025

Exploring Disentangled and Controllable Human Image Synthesis: From End-to-End to Stage-by-Stage

Zhengwentai Sun, Chenghong Li, Hongjie Liao +7

Achieving fine-grained controllability in human image synthesis is a long-standing challenge in computer vision. Existing methods primarily focus on either facial synthesis or near…

cs.CV20241 cited

MVImgNet2.0: A Larger-scale Dataset of Multi-view Images

Xiaoguang Han, Yushuang Wu, Luyue Shi +7

MVImgNet is a large-scale dataset that contains multi-view images of ~220k real-world objects in 238 classes. As a counterpart of ImageNet, it introduces 3D visual signals via mult…

cs.CV2023

MVHumanNet: A Large-scale Dataset of Multi-view Daily Dressing Human Captures

Zhangyang Xiong, Chenghong Li, Kenkun Liu +9

In this era, the success of large language models and text-to-image models can be attributed to the driving force of large-scale datasets. However, in the realm of 3D vision, while…