2 citations · 2 across the 3 of their papers we have counts for
1 paper · 1 filter
Haoyu Wang, Hao Tang, Donglin Di +5
Existing video generation models predominantly emphasize appearance fidelity while exhibiting limited ability to synthesize complex human motions, such as whole-body movements, lon…