2 papers
cs.CL2026
ERNIE 5.0 Technical Report
Haifeng Wang, Hua Wu, Tian Wu +432
In this report, we introduce ERNIE 5.0, a natively autoregressive foundation model desinged for unified multimodal understanding and generation across text, image, video, and audio…
cs.CV2026
A Text-to-3D Framework for Joint Generation of CG-Ready Humans and Compatible Garments
Zhiyao Sun, Yu-Hui Wen, Ho-Jui Fang +4
Creating detailed 3D human avatars with fitted garments traditionally requires specialized expertise and labor-intensive workflows. While recent advances in generative AI have enab…