Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Fed3D: Federated 3D Object Detection
Suyan Dai, Chenxi Liu, Fazeng Li +1
3D object detection models trained in one server plays an important role in autonomous driving, robotics manipulation, and augmented reality scenarios. However, most existing metho…
cs.CV2026
PixelVLA: Advancing Pixel-level Understanding in Vision-Language-Action Model
Wenqi Liang, Gan Sun, Yao He +5
Vision-Language-Action models (VLAs) are emerging as powerful tools for learning generalizable visuomotor control policies. However, current VLAs are mostly trained on large-scale…