2 papers
cs.CV2026
SpatialFusion: Endowing Unified Image Generation with Intrinsic 3D Geometric Awareness
Haiyi Qiu, Kaihang Pan, Jiacheng Li +3
Recent unified image generation models have achieved remarkable success by employing MLLMs for semantic understanding and diffusion backbones for image generation. However, these m…
cs.CV2025
OmniMoGen: Unifying Human Motion Generation via Learning from Interleaved Text-Motion Instructions
Wendong Bu, Kaihang Pan, Yuze Lin +6
Large language models (LLMs) have unified diverse linguistic tasks within a single framework, yet such unification remains unexplored in human motion generation. Existing methods a…