2 papers
cs.CV2025
Video Panels for Long Video Understanding
Lars Doorenbos, Federico Spurio, Juergen Gall
Recent Video-Language Models (VLMs) achieve promising results on long-video understanding, but their performance still lags behind that achieved on tasks involving images or short…
cs.CV2025
Skeleton Motion Words for Unsupervised Skeleton-Based Temporal Action Segmentation
Uzay Gökay, Federico Spurio, Dominik R. Bach +1
Current state-of-the-art methods for skeleton-based temporal action segmentation are predominantly supervised and require annotated data, which is expensive to collect. In contrast…