2 papers
cs.RO2026
Whose Is This?: Context-Aware Object Ownership Inference with Uncertainty-Guided Questioning
Saki Hashimoto, Akira Taniguchi, Shoichi Hasegawa +2
Service robots must infer object ownership to correctly interpret instructions such as "bring me my cup." However, ownership is a latent attribute that cannot be directly observed,…
cs.CL2025
Metropolis-Hastings Captioning Game: Knowledge Fusion of Vision Language Models via Decentralized Bayesian Inference
Yuta Matsui, Ryosuke Yamaki, Ryo Ueda +2
We propose the Metropolis-Hastings Captioning Game (MHCG), a method to fuse knowledge of multiple vision-language models (VLMs) by learning from each other. Although existing metho…