1 citations · 2 across the 2 of their papers we have counts for
3 papers
cs.CV2022★ 1 cited
Efficient Movie Scene Detection using State-Space Transformers
Md Mohaiminul Islam, Mahmudul Hasan, Kishan Shamsundar Athrey +2
The ability to distinguish between different movie scenes is critical for understanding the storyline of a movie. However, accurately detecting movie scenes is often challenging as…
cs.CV2022★ 1 cited
Object State Change Classification in Egocentric Videos using the Divided Space-Time Attention Mechanism
Md Mohaiminul Islam, Gedas Bertasius
This report describes our submission called "TarHeels" for the Ego4D: Object State Change Classification Challenge. We use a transformer-based video recognition model and leverage…
cs.CV2022
Long Movie Clip Classification with State-Space Video Models
Md Mohaiminul Islam, Gedas Bertasius
Most modern video recognition models are designed to operate on short video clips (e.g., 5-10s in length). Thus, it is challenging to apply such models to long movie understanding…