29 citations · 41 across the 3 of their papers we have counts for
1 paper · 1 filter
Georg Heigold, Matthias Minderer, Alexey Gritsenko +5
We present an architecture and a training recipe that adapts pre-trained open-world image models to localization in videos. Understanding the open visual world (without being const…