3 papers
cs.CV2024
Aligning Object Detector Bounding Boxes with Human Preference
Ombretta Strafforello, Osman S. Kayhan, Oana Inel +2
Previous work shows that humans tend to prefer large bounding boxes over small bounding boxes with the same IoU. However, we show here that commonly used object detectors predict l…
cs.CV2023
Video BagNet: short temporal receptive fields increase robustness in long-term action recognition
Ombretta Strafforello, Xin Liu, Klamer Schutte +1
Previous work on long-term video action recognition relies on deep 3D-convolutional models that have a large temporal receptive field (RF). We argue that these models are not alway…
cs.CV2023
Are current long-term video understanding datasets long-term?
Ombretta Strafforello, Klamer Schutte, Jan van Gemert
Many real-world applications, from sport analysis to surveillance, benefit from automatic long-term action recognition. In the current deep learning paradigm for automatic action r…