Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
Superpowering Open-Vocabulary Object Detectors for X-ray Vision
Pablo Garcia-Fernandez, Lorenzo Vaquero, Mingxuan Liu +5
Open-vocabulary object detection (OvOD) is set to revolutionize security screening by enabling systems to recognize any item in X-ray scans. However, developing effective OvOD mode…
cs.CV2024
Lost in Time: A New Temporal Benchmark for VideoLLMs
Daniel Cores, Michael Dorkenwald, Manuel Mucientes +2
Large language models have demonstrated impressive performance when integrated with vision models even enabling video understanding. However, evaluating video models presents its o…
cs.CV2020
Spatio-temporal Tubelet Feature Aggregation and Object Linking in Videos
Daniel Cores, Víctor M. Brea, Manuel Mucientes
This paper addresses the problem of how to exploit spatio-temporal information available in videos to improve the object detection precision. We propose a two stage object detector…