activity
20202022
most citedGraph-MLP: Node Classification without Message Passing in Graph

50 citations · 70 across the 4 of their papers we have counts for

collaborators

5 papers

cs.CV20228 cited

Multimodal Adaptive Distillation for Leveraging Unimodal Encoders for Vision-Language Tasks

Zhecan Wang, Noel Codella, Yen-Chun Chen +8

Cross-modal encoders for vision-language (VL) tasks are often pretrained with carefully curated vision-language datasets. While these datasets reach an order of 10 million samples,…

cs.LG202150 cited

Graph-MLP: Node Classification without Message Passing in Graph

Yang Hu, Haoxuan You, Zhecan Wang +3

Graph Neural Network (GNN) has been demonstrated its effectiveness in dealing with non-Euclidean structural data. Both spatial-based and spectral-based GNNs are relying on adjacenc…

cs.CL2020

Unsupervised Vision-and-Language Pre-training Without Parallel Images and Captions

Liunian Harold Li, Haoxuan You, Zhecan Wang +3

Pre-trained contextual vision-and-language (V&L) models have achieved impressive performance on various benchmarks. However, existing models require a large amount of parallel imag…

cs.CV202010 cited

Learning Visual Commonsense for Robust Scene Graph Generation

Alireza Zareian, Zhecan Wang, Haoxuan You +1

Scene graph generation models understand the scene through object and predicate recognition, but are prone to mistakes due to the challenges of perception in the wild. Perception e…

cs.CV20202 cited

Learning to Detect Head Movement in Unconstrained Remote Gaze Estimation in the Wild

Zhecan Wang, Jian Zhao, Cheng Lu +4

Unconstrained remote gaze estimation remains challenging mostly due to its vulnerability to the large variability in head-pose. Prior solutions struggle to maintain reliable accura…