action tokenization 1behavior cloning 1imitation learning 1motion mode conditioning 1robot manipulation 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.RO2026
MoMo: Dial Motion Mode in Robot Manipulation with Spatiotemporal Action Tokenization
Yuhan Hu, Hugues Thomas, Peide Huang +3
MoMo is a two-stage imitation‑learning system that uses spatiotemporal action tokenization and a behavior‑cloning transformer to let robots vary their manipulation style (steady, d…
cs.CV2025
Pts3D-LLM: Studying the Impact of Token Structure for 3D Scene Understanding With Large Language Models
Hugues Thomas, Chen Chen, Jian Zhang
Effectively representing 3D scenes for Multimodal Large Language Models (MLLMs) is crucial yet challenging. Existing approaches commonly only rely on 2D image features and use vari…
cs.RO2025
DR-MPC: Deep Residual Model Predictive Control for Real-world Social Navigation
James R. Han, Hugues Thomas, Jian Zhang +2
How can a robot safely navigate around people with complex motion patterns? Deep Reinforcement Learning (DRL) in simulation holds some promise, but much prior work relies on simula…