From the 2 of 6 linked papers with an AI index.
7 papers
An LLM-Based Automatic Sportscast Solution for Robot Soccer Matches
Francesco Petri, Michele Brienza, Daniele Nardi +3
The paper presents an autonomous system that extracts statistics from RoboCup soccer video streams and generates real-time, hallucination‑free commentary using a neuro‑symbolic arc…
Semantic Audio-driven Understanding for Dynamic Humanoid Whole Body Control
J. M. A. Marcelo, M. Brienza, E. Bugli +4
The paper presents a framework that enables a humanoid robot to autonomously select and execute whole‑body motion skills in real time based on continuous audio input, using music f…
R2F: Repurposing Ray Frontiers for LLM-free Object Navigation
Francesco Argenziano, John Mark Alexis Marcelo, Michele Brienza +5
Zero-shot open-vocabulary object navigation has progressed rapidly with the emergence of large Vision-Language Models (VLMs) and Large Language Models (LLMs), now widely used as hi…
Context Matters! Relaxing Goals with LLMs for Feasible 3D Scene Planning
Emanuele Musumeci, Michele Brienza, Francesco Argenziano +4
Embodied agents need to plan and act reliably in real and complex 3D environments. Classical planning (e.g., PDDL) offers structure and guarantees, but in practice it fails under n…
LOST-3DSG: Lightweight Open-Vocabulary 3D Scene Graphs with Semantic Tracking in Dynamic Environments
Sara Micol Ferraina, Michele Brienza, Francesco Argenziano +4
Tracking objects that move within dynamic environments is a core challenge in robotics. Recent research has advanced this topic significantly; however, many existing approaches rem…
Multi-Agent Planning Using Visual Language Models
Michele Brienza, Francesco Argenziano, Vincenzo Suriani +2
Large Language Models (LLMs) and Visual Language Models (VLMs) are attracting increasing interest due to their improving performance and applications across various domains and tas…