1 citations · 2 across the 3 of their papers we have counts for
4 papers
SpatialPrompting: Keyframe-driven Zero-Shot Spatial Reasoning with Off-the-Shelf Multimodal Large Language Models
Shun Taguchi, Hideki Deguchi, Takumi Hamazaki +1
This study introduces SpatialPrompting, a novel framework that harnesses the emergent reasoning capabilities of off-the-shelf multimodal large language models to achieve zero-shot…
Online Embedding Multi-Scale CLIP Features into 3D Maps
Shun Taguchi, Hideki Deguchi
This study introduces a novel approach to online embedding of multi-scale CLIP (Contrastive Language-Image Pre-Training) features into 3D maps. By harnessing CLIP, this methodology…
Language to Map: Topological map generation from natural language path instructions
Hideki Deguchi, Kazuki Shibata, Shun Taguchi
In this paper, a method for generating a map from path information described using natural language (textual path) is proposed. In recent years, robotics research mainly focus on v…
CLIP feature-based randomized control using images and text for multiple tasks and robots
Kazuki Shibata, Hideki Deguchi, Shun Taguchi
This study presents a control framework leveraging vision language models (VLMs) for multiple tasks and robots. Notably, existing control methods using VLMs have achieved high perf…