From the 1 of 11 linked papers with an AI index.
11 papers
Think When It Matters: Conditional VLM Reasoning for Social Navigation with RL Policies
Ali Ahmadi, Hamed Rahimi, Adrien Jacquet Cretides +3
The paper presents HUMA, a hybrid system that combines a reactive reinforcement learning navigation policy with a conditional vision-language model to improve semantic and social r…
PRISM: Perception Reasoning Interleaved for Sequential Decision Making
Mohamed Salim Aissi, Clemence Grislain, Clement Romac +4
Scaling LLM-based embodied agents from text-only environments to complex multimodal settings remains a major challenge. Recent work identifies a perception-reasoning-decision gap i…
Encoding Predictability and Legibility for Style-Conditioned Diffusion Policy
Adrien Jacquet Crétides, Mouad Abrini, Hamed Rahimi +1
Striking a balance between efficiency and transparent motion is a core challenge in human-robot collaboration, as highly expressive movements often incur unnecessary time and energ…
IntentVLM: Open-Vocabulary Intention Recognition through Forward-Inverse Modeling with Video-Language Models
Hamed Rahimi, Clemence Grislain, Adrien Jacquet Cretides +2
Improving the effectiveness of human-robot interaction requires social robots to accurately infer human goals through robust intention understanding. This challenge is particularly…
CLUE: Crossmodal disambiguation via Language-vision Understanding with attEntion
Mouad Abrini, Mohamed Chetouani
With the increasing integration of robots into daily life, human-robot interaction has become more complex and multifaceted. A critical component of this interaction is Interactive…
HARMONI: Multimodal Personalization of Multi-User Human-Robot Interactions with LLMs
Jeanne Malécot, Hamed Rahimi, Jeanne Cattoni +5
Existing human-robot interaction systems often lack mechanisms for sustained personalization and dynamic adaptation in multi-user environments, limiting their effectiveness in real…