21 citations · 21 across the 8 of their papers we have counts for
12 papers · 1 filter
OGScene3D: Incremental Open-Vocabulary 3D Gaussian Scene Graph Mapping for Scene Understanding
Siting Zhu, Ziyun Lu, Guangming Wang +5
Open-vocabulary scene understanding is crucial for robotic applications, enabling robots to comprehend complex 3D environmental contexts and supporting various downstream tasks suc…
Rewarding DINO: Predicting Dense Rewards with Vision Foundation Models
Pierre Krack, Tobias Jülg, Wolfram Burgard +1
Well-designed dense reward functions in robot manipulation not only indicate whether a task is completed but also encode progress along the way. Generally, designing dense rewards…
Enabling Dynamic Tracking in Vision-Language-Action Models via Time-Discrete and Time-Continuous Velocity Feedforward
Johannes Hechtl, Philipp Schmitt, Georg von Wichert +1
While vision-language-action (VLA) models have shown great promise for robot manipulation, their deployment on rigid industrial robots remains challenging due to the inherent trade…
VLAgents: A Policy Server for Efficient VLA Inference
Tobias Jülg, Khaled Gamal, Nisarga Nilavadi +5
The rapid emergence of Vision-Language-Action models (VLAs) has a significant impact on robotics. However, their deployment remains complex due to the fragmented interfaces and the…
ConPoSe: LLM-Guided Contact Point Selection for Scalable Cooperative Object Pushing
Noah Steinkrüger, Nisarga Nilavadi, Wolfram Burgard +1
Object transportation in cluttered environments is a fundamental task in various domains, including domestic service and warehouse logistics. In cooperative object transport, multi…
Robot Control Stack: A Lean Ecosystem for Robot Learning at Scale
Tobias Jülg, Pierre Krack, Seongjin Bien +7
Vision-Language-Action models (VLAs) mark a major shift in robot learning. They replace specialized architectures and task-tailored components of expert policies with large-scale d…