1 paper
Akshita Gupta, Aditya Arora, Sanath Narayan +3
Open-Vocabulary Temporal Action Localization (OVTAL) enables a model to recognize any desired action category in videos without the need to explicitly curate training data for all…