1 citations · 2 across the 3 of their papers we have counts for
3 papers
cs.CV2025★ 1 cited
SpatialPrompting: Keyframe-driven Zero-Shot Spatial Reasoning with Off-the-Shelf Multimodal Large Language Models
Shun Taguchi, Hideki Deguchi, Takumi Hamazaki +1
This study introduces SpatialPrompting, a novel framework that harnesses the emergent reasoning capabilities of off-the-shelf multimodal large language models to achieve zero-shot…
cs.RO2024★ 1 cited
Online Embedding Multi-Scale CLIP Features into 3D Maps
Shun Taguchi, Hideki Deguchi
This study introduces a novel approach to online embedding of multi-scale CLIP (Contrastive Language-Image Pre-Training) features into 3D maps. By harnessing CLIP, this methodology…
cs.RO2024
CLIP feature-based randomized control using images and text for multiple tasks and robots
Kazuki Shibata, Hideki Deguchi, Shun Taguchi
This study presents a control framework leveraging vision language models (VLMs) for multiple tasks and robots. Notably, existing control methods using VLMs have achieved high perf…