1 paper · 1 filter
Shoya Kuno, Yumo Ouchi, Kanata Suzuki
Vision-language-action (VLA) models are promising for diverse robotic tasks, but their performance heavily depends on large-scale high-quality training data, whose collection on re…