activity
20162024
most citedRT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

273 citations · 507 across the 16 of their papers we have counts for

collaborators
Showing 2020Show all

7 papers · 1 filter

cs.CV20201 cited

Reducing Inference Latency with Concurrent Architectures for Image Recognition

Ramyad Hadidi, Jiashen Cao, Michael S. Ryoo +1

Satisfying the high computation demand of modern deep learning architectures is challenging for achieving low inference latency. The current approaches in decreasing latency only i…

cs.CV20201 cited

AssembleNet++: Assembling Modality Representations via Attention Connections

Michael S. Ryoo, AJ Piergiovanni, Juhana Kangaspunta +1

We create a family of powerful video models which are able to: (i) learn interactions between semantic object information and raw appearance and motion features, and (ii) deploy at…

cs.CV2020

Adversarial Generative Grammars for Human Activity Prediction

AJ Piergiovanni, Anelia Angelova, Alexander Toshev +1

In this paper we propose an adversarial generative grammar model for future prediction. The objective is to learn a model that explicitly captures temporal dependencies, providing…

cs.CV20207 cited

AttentionNAS: Spatiotemporal Attention Cell Search for Video Classification

Xiaofang Wang, Xuehan Xiong, Maxim Neumann +5

Convolutional operations have two limitations: (1) do not explicitly model where to focus as the same filter is applied to all the positions, and (2) are unsuitable for modeling lo…

cs.CV2020

AViD Dataset: Anonymized Videos from Diverse Countries

AJ Piergiovanni, Michael S. Ryoo

We introduce a new public video dataset for action recognition: Anonymized Videos from Diverse countries (AViD). Unlike existing public video datasets, AViD is a collection of acti…

eess.SP2020

LCP: A Low-Communication Parallelization Method for Fast Neural Network Inference in Image Recognition

Ramyad Hadidi, Bahar Asgari, Jiashen Cao +6

Deep neural networks (DNNs) have inspired new studies in myriad edge applications with robots, autonomous agents, and Internet-of-things (IoT) devices. However, performing inferenc…