From the 2 of 9 linked papers with an AI index.
6 papers · 1 filter
Semantic Audio-driven Understanding for Dynamic Humanoid Whole Body Control
J. M. A. Marcelo, M. Brienza, E. Bugli +4
The paper presents a framework that enables a humanoid robot to autonomously select and execute whole‑body motion skills in real time based on continuous audio input, using music f…
R2F: Repurposing Ray Frontiers for LLM-free Object Navigation
Francesco Argenziano, John Mark Alexis Marcelo, Michele Brienza +5
Zero-shot open-vocabulary object navigation has progressed rapidly with the emergence of large Vision-Language Models (VLMs) and Large Language Models (LLMs), now widely used as hi…
Context Matters! Relaxing Goals with LLMs for Feasible 3D Scene Planning
Emanuele Musumeci, Michele Brienza, Francesco Argenziano +4
Embodied agents need to plan and act reliably in real and complex 3D environments. Classical planning (e.g., PDDL) offers structure and guarantees, but in practice it fails under n…
LOST-3DSG: Lightweight Open-Vocabulary 3D Scene Graphs with Semantic Tracking in Dynamic Environments
Sara Micol Ferraina, Michele Brienza, Francesco Argenziano +4
Tracking objects that move within dynamic environments is a core challenge in robotics. Recent research has advanced this topic significantly; however, many existing approaches rem…
EMPOWER: Embodied Multi-role Open-vocabulary Planning with Online Grounding and Execution
Francesco Argenziano, Michele Brienza, Vincenzo Suriani +2
Task planning for robots in real-life settings presents significant challenges. These challenges stem from three primary issues: the difficulty in identifying grounded sequences of…
LLCoach: Generating Robot Soccer Plans using Multi-Role Large Language Models
Michele Brienza, Emanuele Musumeci, Vincenzo Suriani +4
The deployment of robots into human scenarios necessitates advanced planning strategies, particularly when we ask robots to operate in dynamic, unstructured environments. RoboCup o…