3 papers
cs.RO2024
RT-Grasp: Reasoning Tuning Robotic Grasping via Multi-modal Large Language Model
Jinxuan Xu, Shiyu Jin, Yutian Lei +2
Recent advances in Large Language Models (LLMs) have showcased their remarkable reasoning capabilities, making them influential across various fields. However, in robotics, their u…
cs.RO2024
VIHE: Virtual In-Hand Eye Transformer for 3D Robotic Manipulation
Weiyao Wang, Yutian Lei, Shiyu Jin +2
In this work, we introduce the Virtual In-Hand Eye Transformer (VIHE), a novel method designed to enhance 3D manipulation capabilities through action-aware view rendering. VIHE aut…
cs.RO2024
RLingua: Improving Reinforcement Learning Sample Efficiency in Robotic Manipulations With Large Language Models
Liangliang Chen, Yutian Lei, Shiyu Jin +2
Reinforcement learning (RL) has demonstrated its capability in solving various tasks but is notorious for its low sample efficiency. In this paper, we propose RLingua, a framework…