4 papers
FiCA: Feed-forward instant Gaussian Codec Avatars from a Single Portrait Image
Kim Youwang, Zhengyu Yang, Liuhao Ge +8
We introduce FiCA, a Feed-forward, instant Gaussian Codec Avatar generation pipeline that creates lifelike avatars from a single portrait image. Generating a photorealistic and dri…
MOSLIM:Align with diverse preferences in prompts through reward classification
Yu Zhang, Wanli Jiang, Zhengyu Yang
The multi-objective alignment of Large Language Models (LLMs) is essential for ensuring foundational models conform to diverse human preferences. Current research in this field typ…
Design2Code: Benchmarking Multimodal Code Generation for Automated Front-End Engineering
Chenglei Si, Yanzhe Zhang, Ryan Li +3
Generative AI has made rapid advancements in recent years, achieving unprecedented capabilities in multimodal understanding and code generation. This can enable a new paradigm of f…
GenCA: A Text-conditioned Generative Model for Realistic and Drivable Codec Avatars
Keqiang Sun, Amin Jourabloo, Riddhish Bhalodia +9
Photo-realistic and controllable 3D avatars are crucial for various applications such as virtual and mixed reality (VR/MR), telepresence, gaming, and film production. Traditional m…