2 papers
cs.CV2026
MVOFormer: Flow-Semantic Transformer for Robust Monocular Visual Odometry
Jituo Li, Shunwang Sun, Jialu Zhang +5
Monocular visual odometry (MVO) is foundational to autonomous navigation and robotic localization. However, existing learning-based MVO approaches often struggle with either a lack…
cs.GR2026
DrawVideo: Generating Long Video from Storyboard Keyframe Sketches
Chuanzhi Xu, Huiqi Liang, Bang Shi +7
Long video generation requires high-fidelity synthesis, coherent narrative structure, and user control over extended time spans. Existing text-to-video methods often rely on a sing…