1 citations · 1 across the 16 of their papers we have counts for
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
TABLET: A Large-Scale Dataset for Robust Visual Table Understanding
Iñigo Alonso, Imanol Miranda, Eneko Agirre +1
While table understanding increasingly relies on pixel-only settings, current benchmarks predominantly use synthetic renderings that lack the complexity and visual diversity of rea…
cs.CV2025
Parameter-free Video Segmentation for Vision and Language Understanding
Louis Mahon, Mirella Lapata
The proliferation of creative video content has driven demand for adapting language models to handle video input and enable multimodal understanding. However, end-to-end models str…
cs.CV2024
Finding the Right Moment: Human-Assisted Trailer Creation via Task Composition
Pinelopi Papalampidi, Frank Keller, Mirella Lapata
Movie trailers perform multiple functions: they introduce viewers to the story, convey the mood and artistic style of the film, and encourage audiences to see the movie. These dive…