1 citations · 1 across the 4 of their papers we have counts for
5 papers · 1 filter
Towards Learning a Generalizable 3D Scene Representation from 2D Observations
Martin Gromniak, Jan-Gerrit Habekost, Sebastian Kamp +2
We introduce a Generalizable Neural Radiance Field approach for predicting 3D workspace occupancy from egocentric robot observations. Unlike prior methods operating in camera-centr…
Unified Dynamic Scanpath Predictors Outperform Individually Trained Neural Models
Fares Abawi, Di Fu, Stefan Wermter
Previous research on scanpath prediction has mainly focused on group models, disregarding the fact that the scanpaths and attentional behaviors of individuals are diverse. The disr…
Balancing long- and short-term dynamics for the modeling of saliency in videos
Theodor Wulff, Fares Abawi, Philipp Allgeuer +1
The role of long- and short-term dynamics towards salient object detection in videos is under-researched. We present a Transformer-based approach to learn a joint representation of…
Unconstrained Open Vocabulary Image Classification: Zero-Shot Transfer from Text to Image via CLIP Inversion
Philipp Allgeuer, Kyra Ahrens, Stefan Wermter
We introduce NOVIC, an innovative real-time uNconstrained Open Vocabulary Image Classifier that uses an autoregressive transformer to generatively output classification labels as l…
Concept-Based Explanations in Computer Vision: Where Are We and Where Could We Go?
Jae Hee Lee, Georgii Mikriukov, Gesina Schwalbe +2
Concept-based XAI (C-XAI) approaches to explaining neural vision models are a promising field of research, since explanations that refer to concepts (i.e., semantically meaningful…