3 papers
cs.CL2026
LaViSA: A Language and Vision Structural Ambiguity Benchmark
Lee Sangmyeong, Shun Inadumi, Koichiro Yoshino
Structural ambiguity arises when a single sentence admits multiple valid interpretations due to its syntactic structure, posing a fundamental challenge for language understanding.…
cs.CV2026
SciPostGen: Bridging the Gap between Scientific Papers and Poster Layouts
Shun Inadumi, Shohei Tanaka, Tosho Hirasawa +3
As the number of scientific papers continues to grow, there is a demand for approaches that can effectively convey research findings, with posters serving as a key medium for prese…
cs.CL2025
Disambiguating Reference in Visually Grounded Dialogues through Joint Modeling of Textual and Multimodal Semantic Structures
Shun Inadumi, Nobuhiro Ueda, Koichiro Yoshino
Multimodal reference resolution, including phrase grounding, aims to understand the semantic relations between mentions and real-world objects. Phrase grounding between images and…