4 citations · 6 across the 4 of their papers we have counts for
1 paper · 1 filter
Savas Ozkan, Gozde Bozdagi Akar
Frame-level visual features are generally aggregated in time with the techniques such as LSTM, Fisher Vectors, NetVLAD etc. to produce a robust video-level representation. We here…