3 papers
cs.RO2026
VANE: Reliable Test-Time Training for Vision-Language-Action Models via Future Visual Representation Prediction
Hongjin Ji, Guoyang Xia, Luoyang Sun +2
Test-time training (TTT) offers a lightweight way to adapt vision--language--action (VLA) policies from unlabeled deployment streams, but it remains difficult to use reliably in cl…
cs.CV2026
VLAFlow: A Unified Training Framework for Vision-Language-Action Models via Co-training and Future Latent Alignment
Guoyang Xia, Fengfa Li, Hongjin Ji +4
Vision-language-action models (VLAs) have recently advanced robotic manipulation, yet the effects of different robot-data pre-training paradigms remain difficult to compare because…
math.ST2026
Optimal Estimation in Orthogonally Invariant Generalized Linear Models: Spectral Initialization and Approximate Message Passing
Yihan Zhang, Hong Chang Ji, Ramji Venkataramanan +1
We consider the problem of parameter estimation from a generalized linear model with a random design matrix that is orthogonally invariant in law. Such a model allows the design ha…