3 papers
cs.IR2026
Same Image, Different Meanings: Toward Retrieval of Context-Dependent Meanings
Ayuto Tsutsumi, Ryosuke Kohita
A scene of two people in the rain can convey hope and warmth in a reunion story or sorrow and finality in a farewell story. We investigate this context-dependent nature of image me…
cs.SD2026
The TMU System for the XACLE Challenge: Training Large Audio Language Models with CLAP Pseudo-Labels
Ayuto Tsutsumi, Kohei Tanaka, Sayaka Shiota
In this paper, we propose a submission to the x-to-audio alignment (XACLE) challenge. The goal is to predict semantic alignment of a given general audio and text pair. The proposed…
cs.CL2025
Do Large Language Models Know Folktales? A Case Study of Yokai in Japanese Folktales
Ayuto Tsutsumi, Yuu Jinnai
Although Large Language Models (LLMs) have demonstrated strong language understanding and generation abilities across various languages, their cultural knowledge is often limited t…