101 citations · 119 across the 8 of their papers we have counts for
24 papers
AlignVE: Visual Entailment Recognition Based on Alignment Relations
Biwei Cao, Jiuxin Cao, Jie Gui +5
Visual entailment (VE) is to recognize whether the semantics of a hypothesis text can be inferred from the given premise image, which is one special task among recent emerged visio…
Searching a High-Performance Feature Extractor for Text Recognition Network
Hui Zhang, Quanming Yao, James T. Kwok +1
Feature extractor plays a critical role in text recognition (TR), but customizing its architecture is relatively less explored due to expensive manual tweaking. In this work, inspi…
Low-rank Tensor Learning with Nonconvex Overlapped Nuclear Norm Regularization
Quanming Yao, Yaqing Wang, Bo Han +1
Nonconvex regularization has been popularly used in low-rank matrix learning. However, extending it for low-rank tensor learning is still computationally expensive. To address this…
Revisiting Over-smoothing in BERT from the Perspective of Graph
Han Shi, Jiahui Gao, Hang Xu +5
Recently over-smoothing phenomenon of Transformer-based models is observed in both vision and language fields. However, no existing work has delved deeper to further investigate th…
Dropout's Dream Land: Generalization from Learned Simulators to Reality
Zac Wellmer, James T. Kwok
A World Model is a generative model used to simulate an environment. World Models have proven capable of learning spatial and temporal representations of Reinforcement Learning env…
Pyramidal Dense Attention Networks for Lightweight Image Super-Resolution
Huapeng Wu, Jie Gui, Jun Zhang +2
Recently, deep convolutional neural network methods have achieved an excellent performance in image superresolution (SR), but they can not be easily applied to embedded devices due…