2 papers
cs.CV2026
PREGEN: Uncovering Latent Thoughts in Composed Video Retrieval
Gabriele Serussi, David Vainshtein, Jonathan Kouchly +2
Composed Video Retrieval (CoVR) aims to retrieve a video based on a query video and a modifying text. Current CoVR methods fail to fully exploit modern Vision-Language Models (VLMs…
cs.CV2025
Robot Instance Segmentation with Few Annotations for Grasping
Moshe Kimhi, David Vainshtein, Chaim Baskin +1
The ability of robots to manipulate objects relies heavily on their aptitude for visual perception. In domains characterized by cluttered scenes and high object variability, most m…