Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Scene-VLM: Multimodal Video Scene Segmentation via Vision-Language Models
Nimrod Berman, Adam Botach, Emanuel Ben-Baruch +5
Segmenting long-form videos into semantically coherent scenes is a fundamental task in large-scale video understanding. Existing encoder-based methods are limited by visual-centric…
cs.CV2025
Tell Me What You See: Text-Guided Real-World Image Denoising
Erez Yosef, Raja Giryes
Image reconstruction from noisy sensor measurements is challenging and many methods have been proposed for it. Yet, most approaches focus on learning robust natural image priors wh…