2 papers
cs.RO2026
WALL-WM: Carving World Action Modeling at the Event Joints
Shalfun Li, Victor Yao, Charles Yang +28
WALL-WM is a World Action Model that shifts video-action learning from chunk-centric optimization to event-grounded Vision-Language-Action pretraining, using semantically coherent…
cs.CV2025
ERUPT: Efficient Rendering with Unposed Patch Transformer
Maxim V. Shugaev, Vincent Chen, Maxim Karrenbach +3
This work addresses the problem of novel view synthesis in diverse scenes from small collections of RGB images. We propose ERUPT (Efficient Rendering with Unposed Patch Transformer…