2 papers
cs.CV2021
Exploration of Visual Features and their weighted-additive fusion for Video Captioning
Praveen S, Akhilesh Bharadwaj, Harsh Raj +4
Video captioning is a popular task that challenges models to describe events in videos using natural language. In this work, we investigate the ability of various visual feature re…
cs.CV2020
Knowledge Fusion Transformers for Video Action Recognition
Ganesh Samarth, Sheetal Ojha, Nikhil Pareek
We introduce Knowledge Fusion Transformers for video action classification. We present a self-attention based feature enhancer to fuse action knowledge in 3D inception based spatio…