most citedGETAvatar: Generative Textured Meshes for Animatable Human Avatars

1 citations · 1 across the 3 of their papers we have counts for

collaborators

5 papers

cs.SD2026

FoleyDirector: Fine-Grained Temporal Steering for Video-to-Audio Generation via Structured Scripts

You Li, Dewei Zhou, Fan Ma +3

Recent Video-to-Audio (V2A) methods have achieved remarkable progress, enabling the synthesis of realistic, high-quality audio. However, they struggle with fine-grained temporal co…

cs.CV2026

PhyEdit: Towards Real-World Object Manipulation via Physically-Grounded Image Editing

Ruihang Xu, Dewei Zhou, Xiaolong Shen +2

Achieving physically accurate object manipulation in image editing is essential for its potential applications in interactive world models. However, existing visual generative mode…

cs.CV2023

Vista-LLaMA: Reducing Hallucination in Video Language Models via Equal Distance to Visual Tokens

Fan Ma, Xiaojie Jin, Heng Wang +3

Recent advances in large video-language models have displayed promising outcomes in video comprehension. Current approaches straightforwardly convert video into language tokens and…

cs.GR2023

AvatarStudio: High-fidelity and Animatable 3D Avatar Creation from Text

Jianfeng Zhang, Xuanmeng Zhang, Huichao Zhang +4

We study the problem of creating high-fidelity and animatable 3D avatars from only textual descriptions. Existing text-to-avatar methods are either limited to static avatars which…

cs.CV20231 cited

GETAvatar: Generative Textured Meshes for Animatable Human Avatars

Xuanmeng Zhang, Jianfeng Zhang, Rohan Chacko +4

We study the problem of 3D-aware full-body human generation, aiming at creating animatable human avatars with high-quality textures and geometries. Generally, two challenges remain…