2 papers
cs.CV2026
Scene-VLM: Multimodal Video Scene Segmentation via Vision-Language Models
Nimrod Berman, Adam Botach, Emanuel Ben-Baruch +5
Segmenting long-form videos into semantically coherent scenes is a fundamental task in large-scale video understanding. Existing encoder-based methods are limited by visual-centric…
cs.CL2025
Group-Aware Reinforcement Learning for Output Diversity in Large Language Models
Oron Anschel, Alon Shoshan, Adam Botach +7
Large Language Models (LLMs) often suffer from mode collapse, repeatedly generating the same few completions even when many valid answers exist, limiting their diversity across a w…