Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Vision and Language Reference Prompt into SAM for Few-shot Segmentation
Kosuke Sakurai, Ryotaro Shimizu, Masayuki Goto
Segment Anything Model (SAM) represents a large-scale segmentation model that enables powerful zero-shot capabilities with flexible prompts. While SAM can segment any object in zer…
cs.CV2024
LARE: Latent Augmentation using Regional Embedding with Vision-Language Model
Kosuke Sakurai, Tatsuya Ishii, Ryotaro Shimizu +2
In recent years, considerable research has been conducted on vision-language models that handle both image and text data; these models are being applied to diverse downstream tasks…