5 papers · 1 filter
Evaluation of Multilingual Ability to Use Spatial Deictic Expressions in Vision-Language Models
Kaito Watanabe, Taisei Yamamoto, Tomoki Doi +1
One of the expected abilities of vision-language models (VLMs) is spatial reasoning ability based on a given text and image. To evaluate the spatial reasoning abilities of VLMs, we…
Developing a Guideline for the Labovian-Structural Analysis of Oral Narratives in Japanese
Amane Watahiki, Tomoki Doi, Akari Kikuchi +5
Narrative analysis is a cornerstone of qualitative research. One leading approach is the Labovian model, but its application is labor-intensive, requiring a holistic, recursive int…
Investigating Training and Generalization in Faithful Self-Explanations of Large Language Models
Tomoki Doi, Masaru Isonuma, Hitomi Yanaka
Large language models have the potential to generate explanations for their own predictions in a variety of styles based on user instructions. Recent research has examined whether…
Bridging Perception and Language: A Systematic Benchmark for LVLMs' Understanding of Amodal Completion Reports
Amane Watahiki, Tomoki Doi, Taiga Shinozaki +4
One of the main objectives in developing large vision-language models (LVLMs) is to engineer systems that can assist humans with multimodal tasks, including interpreting descriptio…
Comprehensive Evaluation of Large Language Models for Topic Modeling
Tomoki Doi, Masaru Isonuma, Hitomi Yanaka
Recent work utilizes Large Language Models (LLMs) for topic modeling, generating comprehensible topic labels for given documents. However, their performance has mainly been evaluat…