Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
SCOT: Self-Supervised Contrastive Pretraining For Zero-Shot Compositional Retrieval
Bhavin Jawade, Joao V. B. Soares, Kapil Thadani +6
Compositional image retrieval (CIR) is a multimodal learning task where a model combines a query image with a user-provided text modification to retrieve a target image. CIR finds…
cs.CV2024
Salient Object-Aware Background Generation using Text-Guided Diffusion Models
Amir Erfan Eshratifar, Joao V. B. Soares, Kapil Thadani +4
Generating background scenes for salient objects plays a crucial role across various domains including creative design and e-commerce, as it enhances the presentation and context o…