causal pretraining 1foundation models 1robot control 1sparse mixture of experts 1video-action models 1
From the 1 of 2 linked papers with an AI index.
2 papers
cs.RO2026
Native Video-Action Pretraining for Generalizable Robot Control
Qihang Zhang, Lin Li, Luyao Zhang +26
The paper introduces LingBot-VA 2.0, a video-action foundation model designed specifically for robot control, featuring a semantic visual-action tokenizer, causal pretraining, a sp…
cs.CV2025
LINR Bridge: Vector Graphic Animation via Neural Implicits and Video Diffusion Priors
Wenshuo Gao, Xicheng Lan, Luyao Zhang +1
Vector graphics, known for their scalability and user-friendliness, provide a unique approach to visual content compared to traditional pixel-based images. Animation of these graph…