collaborators
Showing cs.ROShow all

6 papers · 1 filter

cs.RO2024

RT-Grasp: Reasoning Tuning Robotic Grasping via Multi-modal Large Language Model

Jinxuan Xu, Shiyu Jin, Yutian Lei +2

Recent advances in Large Language Models (LLMs) have showcased their remarkable reasoning capabilities, making them influential across various fields. However, in robotics, their u…

cs.RO2024

ExACT: An End-to-End Autonomous Excavator System Using Action Chunking With Transformers

Liangliang Chen, Shiyu Jin, Haoyu Wang +1

Excavators are crucial for diverse tasks such as construction and mining, while autonomous excavator systems enhance safety and efficiency, address labor shortages, and improve hum…

cs.RO2024

VIHE: Virtual In-Hand Eye Transformer for 3D Robotic Manipulation

Weiyao Wang, Yutian Lei, Shiyu Jin +2

In this work, we introduce the Virtual In-Hand Eye Transformer (VIHE), a novel method designed to enhance 3D manipulation capabilities through action-aware view rendering. VIHE aut…

cs.RO2024

RLingua: Improving Reinforcement Learning Sample Efficiency in Robotic Manipulations With Large Language Models

Liangliang Chen, Yutian Lei, Shiyu Jin +2

Reinforcement learning (RL) has demonstrated its capability in solving various tasks but is notorious for its low sample efficiency. In this paper, we propose RLingua, a framework…

cs.RO2024

Reasoning Grasping via Multimodal Large Language Model

Shiyu Jin, Jinxuan Xu, Yutian Lei +1

Despite significant progress in robotic systems for operation within human-centric environments, existing models still heavily rely on explicit human commands to identify and manip…

cs.RO20233 cited

GRAINS: Proximity Sensing of Objects in Granular Materials

Zeqing Zhang, Ruixing Jia, Youcan Yan +5

Proximity sensing detects an object's presence without contact. However, research has rarely explored proximity sensing in granular materials (GM) due to GM's lack of visual and co…