5 citations · 5 across the 4 of their papers we have counts for
4 papers
Video Pre-trained Transformer: A Multimodal Mixture of Pre-trained Experts
Kastan Day, Daniel Christl, Rohan Salvi +1
We present Video Pre-trained Transformer. VPT uses four SOTA encoder models from prior work to convert a video into a sequence of compact embeddings. Our backbone, based on a refer…
CoCAtt: A Cognitive-Conditioned Driver Attention Dataset (Supplementary Material)
Yuan Shen, Niviru Wijayaratne, Pranav Sriram +3
The task of driver attention prediction has drawn considerable interest among researchers in robotics and the autonomous vehicle industry. Driver attention prediction can play an i…
Meta-path Analysis on Spatio-Temporal Graphs for Pedestrian Trajectory Prediction
Aamir Hasan, Pranav Sriram, Katherine Driggs-Campbell
Spatio-temporal graphs (ST-graphs) have been used to model time series tasks such as traffic forecasting, human motion modeling, and action recognition. The high-level structure an…
CoCAtt: A Cognitive-Conditioned Driver Attention Dataset
Yuan Shen, Niviru Wijayaratne, Pranav Sriram +3
The task of driver attention prediction has drawn considerable interest among researchers in robotics and the autonomous vehicle industry. Driver attention prediction can play an i…