4 papers
Temporal Preservation over Processing: Diagnosing and Designing Spatiotemporal Single-Stage Video Detectors
Karam Tomotaki-Dawoud, Anna Hilsmann, Peter Eisert +1
Single-stage video object detectors are increasingly deployed in time-critical applications, yet it remains unclear whether these models genuinely reason over temporal context or m…
Monocular 3D Object Position Estimation with VLMs for Human-Robot Interaction
Ari Wahl, Dorian Gawlinski, David Przewozny +3
Pre-trained general-purpose Vision-Language Models (VLM) hold the potential to enhance intuitive human-machine interactions due to their rich world knowledge and 2D object detectio…
AppleGrowthVision: A large-scale stereo dataset for phenological analysis, fruit detection, and 3D reconstruction in apple orchards
Laura-Sophia von Hirschhausen, Jannes S. Magnusson, Mykyta Kovalenko +6
Deep learning has transformed computer vision for precision agriculture, yet apple orchard monitoring remains limited by dataset constraints. The lack of diverse, realistic dataset…
EEG-Features for Generalized Deepfake Detection
Arian Beckmann, Tilman Stephani, Felix Klotzsche +8
Since the advent of Deepfakes in digital media, the development of robust and reliable detection mechanism is urgently called for. In this study, we explore a novel approach to Dee…