3 papers
cs.CV2025
Decoding Vision Transformers: the Diffusion Steering Lens
Ryota Takatsuki, Sonia Joseph, Ippei Fujisawa +1
Logit Lens is a widely adopted method for mechanistic interpretability of transformer-based language models, enabling the analysis of how internal representations evolve across lay…
cs.CV2025
MCM: Multi-layer Concept Map for Efficient Concept Learning from Masked Images
Yuwei Sun, Lu Mi, Ippei Fujisawa +4
Masking strategies commonly employed in natural language processing are still underexplored in vision tasks such as concept learning, where conventional methods typically rely on f…
cs.AI2024
ProcBench: Benchmark for Multi-Step Reasoning and Following Procedure
Ippei Fujisawa, Sensho Nobe, Hiroki Seto +5
Reasoning is central to a wide range of intellectual activities, and while the capabilities of large language models (LLMs) continue to advance, their performance in reasoning task…