2 papers
cs.CV2026
Archon: A Unified Multimodal Model for Holistic Digital Human Generation
Chong Bao, Shichen Liu, Lijun Yu +9
Digital humans are fundamental to immersive interaction, yet creating a unified model for holistic modalities, including text, audio, motion, and visual content, remains an open ch…
cs.CV2024
A Global Depth-Range-Free Multi-View Stereo Transformer Network with Pose Embedding
Yitong Dong, Yijin Li, Zhaoyang Huang +6
In this paper, we propose a novel multi-view stereo (MVS) framework that gets rid of the depth range prior. Unlike recent prior-free MVS methods that work in a pair-wise manner, ou…