4 papers
MASS: Multiplayer World Models with Authoritative Shared State
Ziqi Cai, Siqi Yang, Yimu Wang +6
Current video world models struggle in multiplayer environments because they entangle world state with view-dependent visual latents, leading to redundant compute, view inconsisten…
Generative World Renderer at the Speed of Play
Guixu Lin, Zheng-Hui Huang, Siqi Yang +3
Generative world renderer AlayaRenderer receives structured world states exported from physics engines and synthesizes RGB frames. Unlike models that generate frames from text/cont…
InstructAV2AV: Instruction-Guided Audio-Video Joint Editing
Haojie Zheng, Yixin Yang, Siqi Yang +2
Recent diffusion-based methods have achieved impressive progress in video content manipulation. However, they typically ignore the accompanying audio, leaving the audio disjointed…
High-Speed Full-Color HDR Imaging via Unwrapping Modulo-Encoded Spike Streams
Chu Zhou, Siqi Yang, Kailong Zhang +4
Conventional RGB-based high dynamic range (HDR) imaging faces a fundamental trade-off between motion artifacts in multi-exposure captures and irreversible information loss in singl…