1 paper · 1 filter
Miranda Muqing Miao, Subin Kim, Brandon Yang +1
Vision-Language-Action (VLA) models leverage powerful perceptual priors from web-scale Vision-Language Model (VLM) pre-training, yet they remain surprisingly brittle in practice, f…