3 papers
cs.CV2026
FactorizedHMR: A Hybrid Framework for Video Human Mesh Recovery
Patrick Kwon, Chen Chen
Human Mesh Recovery (HMR) is fundamentally ambiguous: under occlusion or weak depth cues, multiple 3D bodies can explain the same image evidence. This ambiguity is not uniform acro…
cs.CV2025
DreamingComics: A Story Visualization Pipeline via Subject and Layout Customized Generation using Video Models
Patrick Kwon, Chen Chen
Current story visualization methods tend to position subjects solely by text and face challenges in maintaining artistic consistency. To address these limitations, we introduce Dre…
cs.CV2025
GraspDiffusion: Synthesizing Realistic Whole-body Hand-Object Interaction
Patrick Kwon, Chen Chen, Hanbyul Joo
Recent generative models can synthesize high-quality images, but they often fail to generate humans interacting with objects using their hands. This arises mostly from the model's…