3 papers
cs.CV2026
PinPoint: Prompting with Informative Interior Points
Pouya Sadeghi, Shawn He, Pedro Pablo Guerrero Vela +3
Modern referring image segmentation pipelines couple a vision-language model (VLM) for grounding with a promptable segmenter such as the Segment Anything Model (SAM) for mask gener…
cs.CV2026
Zero-Shot Object Re-Identification in Egocentric Kitchen Videos via Multi-Stage SAM3 Feature Fusion
Dmytro Klepachevskyi, Alexander Wong, Sirisha Rambhatla +1
Object re-identification (ReID) in egocentric kitchen videos is challenging due to rapid viewpoint changes, frequent occlusions, cluttered scenes, and large intra-class appearance…
cs.CV2025
LangDA: Building Context-Awareness via Language for Domain Adaptive Semantic Segmentation
Chang Liu, Bavesh Balaji, Saad Hossain +5
Unsupervised domain adaptation for semantic segmentation (DASS) aims to transfer knowledge from a label-rich source domain to a target domain with no labels. Two key approaches in…