activity
20142023
most citedFaster VoxelPose: Real-time 3D Human Pose Estimation by Orthographic Projection

5 citations · 14 across the 9 of their papers we have counts for

collaborators
Showing cs.CVShow all

9 papers · 1 filter

cs.CV20231 cited

MicroCinema: A Divide-and-Conquer Approach for Text-to-Video Generation

Yanhui Wang, Jianmin Bao, Wenming Weng +12

We present MicroCinema, a straightforward yet effective framework for high-quality and coherent text-to-video generation. Unlike existing approaches that align text prompts with vi…

cs.CV20235 cited

V-DETR: DETR with Vertex Relative Position Encoding for 3D Object Detection

Yichao Shen, Zigang Geng, Yuhui Yuan +6

We introduce a highly performant 3D object detector for point clouds using the DETR framework. The prior attempts all end up with suboptimal results because they fail to learn accu…

cs.CV20231 cited

Human Pose as Compositional Tokens

Zigang Geng, Chunyu Wang, Yixuan Wei +3

Human pose is typically represented by a coordinate vector of body joints or their heatmap embeddings. While easy for data processing, unrealistic pose estimates are admitted due t…

cs.CV2022

Robust Multi-Object Tracking by Marginal Inference

Yifu Zhang, Chunyu Wang, Xinggang Wang +2

Multi-object tracking in videos requires to solve a fundamental problem of one-to-one assignment between objects in adjacent frames. Most methods address the problem by first disca…

cs.CV20225 cited

Faster VoxelPose: Real-time 3D Human Pose Estimation by Orthographic Projection

Hang Ye, Wentao Zhu, Chunyu Wang +2

While the voxel-based methods have achieved promising results for multi-person 3D pose estimation from multi-cameras, they suffer from heavy computation burdens, especially for lar…

cs.CV20221 cited

VirtualPose: Learning Generalizable 3D Human Pose Models from Virtual Data

Jiajun Su, Chunyu Wang, Xiaoxuan Ma +2

While monocular 3D pose estimation seems to have achieved very accurate results on the public datasets, their generalization ability is largely overlooked. In this work, we perform…