3 papers
cs.MA2026
Multi-Agent Robotic Control with Onboard Vision-Language Models
Kajetan Rachwał, Maciej Majek, Bartłomiej Boczek +6
Vision Language Models (VLMs) and Vision Language Action (VLA) models have shown promise in robotic control. Yet, they face significant challenges regarding explainability, general…
cs.MA2025
RAI: Flexible Agent Framework for Embodied AI
Kajetan Rachwał, Maciej Majek, Bartłomiej Boczek +4
With an increase in the capabilities of generative language models, a growing interest in embodied AI has followed. This contribution introduces RAI - a framework for creating embo…
cs.CV2023
Towards Edge-Cloud Architectures for Personal Protective Equipment Detection
Jaroslaw Legierski, Kajetan Rachwal, Piotr Sowinski +5
Detecting Personal Protective Equipment in images and video streams is a relevant problem in ensuring the safety of construction workers. In this contribution, an architecture enab…