activity
20192024
most citedDD-PPO: Learning Near-Perfect PointGoal Navigators from 2.5 Billion Frames

36 citations · 188 across the 35 of their papers we have counts for

collaborators
Showing 2022Show all

11 papers · 1 filter

cs.RO2022

Cascaded Compositional Residual Learning for Complex Interactive Behaviors

K. Niranjan Kumar, Irfan Essa, Sehoon Ha

Real-world autonomous missions often require rich interaction with nearby objects, such as doors or switches, along with effective navigation. However, such complex behaviors are d…

cs.CV2022★ 1 cited

MAGVIT: Masked Generative Video Transformer

Lijun Yu, Yong Cheng, Kihyuk Sohn +8

We introduce the MAsked Generative VIdeo Transformer, MAGVIT, to tackle various video synthesis tasks with a single model. We introduce a 3D tokenizer to quantize a video into spat…

cs.LG2022

Investigating Enhancements to Contrastive Predictive Coding for Human Activity Recognition

Harish Haresamudram, Irfan Essa, Thomas Ploetz

The dichotomy between the challenging nature of obtaining annotations for activities, and the more straightforward nature of data collection from wearables, has resulted in signifi…

cs.CV2022★ 4 cited

Multi-Stage Based Feature Fusion of Multi-Modal Data for Human Activity Recognition

Hyeongju Choi, Apoorva Beedu, Harish Haresamudram +1

To properly assist humans in their needs, human activity recognition (HAR) systems need the ability to fuse information from multiple modalities. Our hypothesis is that multimodal…

cs.CV2022★ 6 cited

Video based Object 6D Pose Estimation using Transformers

Apoorva Beedu, Huda Alamri, Irfan Essa

We introduce a Transformer based 6D Object Pose Estimation framework VideoPose, comprising an end-to-end attention based modelling architecture, that attends to previous frames in…

cs.CV2022★ 1 cited

End-to-End Multimodal Representation Learning for Video Dialog

Huda Alamri, Anthony Bilic, Michael Hu +2

Video-based dialog task is a challenging multimodal learning task that has received increasing attention over the past few years with state-of-the-art obtaining new performance rec…