3 citations · 4 across the 3 of their papers we have counts for
3 papers
Knowing Where to Focus: Event-aware Transformer for Video Grounding
Jinhyun Jang, Jungin Park, Jin Kim +2
Recent DETR-based video grounding models have made the model directly predict moment timestamps without any hand-crafted components, such as a pre-defined proposal or non-maximum s…
Learning to Detect Touches on Cluttered Tables
Norberto Adrian Goussies, Kenji Hata, Shruthi Prabhakara +28
We present a novel self-contained camera-projector tabletop system with a lamp form-factor that brings digital intelligence to our tables. We propose a real-time, on-device, learni…
Probabilistic Prompt Learning for Dense Prediction
Hyeongjun Kwon, Taeyong Song, Somi Jeong +3
Recent progress in deterministic prompt learning has become a promising alternative to various downstream vision tasks, enabling models to learn powerful visual representations wit…