2 papers
cs.RO2025
cVLA: Towards Efficient Camera-Space VLAs
Max Argus, Jelena Bratulic, Houman Masnavi +4
Vision-Language-Action (VLA) models offer a compelling framework for tackling complex robotic manipulation tasks, but they are often expensive to train. In this paper, we propose a…
cs.CV2025
Label-Efficient LiDAR Semantic Segmentation with 2D-3D Vision Transformer Adapters
Julia Hindel, Rohit Mohan, Jelena Bratulic +3
LiDAR semantic segmentation models are typically trained from random initialization as universal pre-training is hindered by the lack of large, diverse datasets. Moreover, most poi…