15 citations · 19 across the 3 of their papers we have counts for
6 papers
Learning Temporal Rules from Noisy Timeseries Data
Karan Samel, Zelin Zhao, Binghong Chen +4
Events across a timeline are a common data representation, seen in different temporal modalities. Individual atomic events can occur in a certain temporal ordering to compose highe…
ProTo: Program-Guided Transformer for Program-Guided Tasks
Zelin Zhao, Karan Samel, Binghong Chen +1
Programs, consisting of semantic and structural information, play an important role in the communication between humans and agents. Towards learning general program executors to un…
How to Design Sample and Computationally Efficient VQA Models
Karan Samel, Zelin Zhao, Binghong Chen +3
In multi-modal reasoning tasks, such as visual question answering (VQA), there have been many modeling and training paradigms tested. Previous models propose different methods for…
Augmenting Policy Learning with Routines Discovered from a Single Demonstration
Zelin Zhao, Chuang Gan, Jiajun Wu +2
Humans can abstract prior knowledge from very little data and use it to boost skill learning. In this paper, we propose routine-augmented policy learning (RAPL), which discovers ro…
Estimating 6D Pose From Localizing Designated Surface Keypoints
Zelin Zhao, Gao Peng, Haoyu Wang +3
In this paper, we present an accurate yet effective solution for 6D pose estimation from an RGB image. The core of our approach is that we first designate a set of surface points o…
PointSIFT: A SIFT-like Network Module for 3D Point Cloud Semantic Segmentation
Mingyang Jiang, Yiran Wu, Tianqi Zhao +2
Recently, 3D understanding research sheds light on extracting features from point cloud directly, which requires effective shape pattern description of point clouds. Inspired by th…