3 papers
cs.CV2026
Towards Vision-Free CIR: Attribute-Augmented Scoring and LLM-Based Reranking for Zero-Shot Composed Image Retrieval
Ryotaro Shimada, Yu-Chieh Lin, Yuji Nozawa +3
Recent work has shown that "Vision-Free'' approaches (representing images as text) can be effective for standard image retrieval tasks. However, it remains unclear whether this par…
cs.CV2025
Prompt-Guided Attention Head Selection for Focus-Oriented Image Retrieval
Yuji Nozawa, Yu-Chieh Lin, Kazumoto Nakamura +1
The goal of this paper is to enhance pretrained Vision Transformer (ViT) models for focus-oriented image retrieval with visual prompting. In real-world image retrieval scenarios, b…
cs.CV2024
Improving Image Clustering with Artifacts Attenuation via Inference-Time Attention Engineering
Kazumoto Nakamura, Yuji Nozawa, Yu-Chieh Lin +2
The goal of this paper is to improve the performance of pretrained Vision Transformer (ViT) models, particularly DINOv2, in image clustering task without requiring re-training or f…