5 citations · 9 across the 16 of their papers we have counts for
4 papers · 1 filter
LumiTokens: 3D Relighting via Token-Space Lighting Transformation
Yiwen Chen, Matheus Gadelha, Huaizu Jiang
Existing 3D relighting methods operate through either explicit material decomposition, diffusion-based view-space generation, or a combination of both, requiring full recomputation…
Surface Keypoint Representation for Multi-Object and Articulated Human-Object Interaction Generation
Xiaogang Peng, Zeyu Han, Zichong Meng +4
Daily activities require humans to coordinate whole-body motion with the motion of surrounding objects. Despite recent progress in human-object interaction (HOI) generation, most e…
Streaming Video Generation with Streaming Force Control
Hanhui Wang, Yiming Xie, Haiwen Feng +3
We introduce StreamForce, a streaming video generation framework that enables physically grounded control through continuous force inputs. Unlike prior video models that train sepa…
SocialFusion: Addressing Social Degradation in Pre-trained Vision-Language Models
Hamza Tahboub, Weiyan Shi, Gang Hua +1
Understanding social interactions from visual cues is a fundamental challenge for a socially competent AI. While powerful pre-trained vision-language models (VLMs) have shown remar…