2 papers
cs.CV2025
Superpowering Open-Vocabulary Object Detectors for X-ray Vision
Pablo Garcia-Fernandez, Lorenzo Vaquero, Mingxuan Liu +5
Open-vocabulary object detection (OvOD) is set to revolutionize security screening by enabling systems to recognize any item in X-ray scans. However, developing effective OvOD mode…
cs.CV2025
Lost in Time: A New Temporal Benchmark for VideoLLMs
Daniel Cores, Michael Dorkenwald, Manuel Mucientes +2
Large language models have demonstrated impressive performance when integrated with vision models even enabling video understanding. However, evaluating video models presents its o…