1 paper
Kazuki Omi, Jion Oshima, Toru Tamaki
This paper proposes a method for spatio-temporal action detection (STAD) that directly generates action tubes from the original video without relying on post-processing steps such…