4 citations · 4 across the 2 of their papers we have counts for
3 papers
cs.CL2022
EfficientVLM: Fast and Accurate Vision-Language Models via Knowledge Distillation and Modal-adaptive Pruning
Tiannan Wang, Wangchunshu Zhou, Yan Zeng +1
Pre-trained vision-language models (VLMs) have achieved impressive results in a range of vision-language tasks. However, popular VLMs usually consist of hundreds of millions of par…
cs.CV2022★ 4 cited
VLUE: A Multi-Task Benchmark for Evaluating Vision-Language Models
Wangchunshu Zhou, Yan Zeng, Shizhe Diao +1
Recent advances in vision-language pre-training (VLP) have demonstrated impressive performance in a range of vision-language (VL) tasks. However, there exist several challenges for…
cs.CL2018
Multi-labeled Relation Extraction with Attentive Capsule Network
Xinsong Zhang, Pengshuai Li, Weijia Jia +1
To disclose overlapped multiple relations from a sentence still keeps challenging. Most current works in terms of neural models inconveniently assuming that each sentence is explic…