34 citations · 160 across the 32 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
TABLET: A Large-Scale Dataset for Robust Visual Table Understanding
Iñigo Alonso, Imanol Miranda, Eneko Agirre +1
While table understanding increasingly relies on pixel-only settings, current benchmarks predominantly use synthetic renderings that lack the complexity and visual diversity of rea…
cs.CV2025
Parameter-free Video Segmentation for Vision and Language Understanding
Louis Mahon, Mirella Lapata
The proliferation of creative video content has driven demand for adapting language models to handle video input and enable multimodal understanding. However, end-to-end models str…