activity
20162024
most citedAn End-to-End Spatio-Temporal Attention Model for Human Action Recognition from Skeleton Data

483 citations · 496 across the 9 of their papers we have counts for

collaborators
Showing 2023Show all

5 papers · 1 filter

cs.CV2023

Shatter and Gather: Learning Referring Image Segmentation with Text Supervision

Dongwon Kim, Namyup Kim, Cuiling Lan +1

Referring image segmentation, the task of segmenting any arbitrary entities described in free-form texts, opens up a variety of vision applications. However, manual labeling of tra…

cs.CV202310 cited

Adaptive Frequency Filters As Efficient Global Token Mixers

Zhipeng Huang, Zhizheng Zhang, Cuiling Lan +3

Recent vision transformers, large-kernel CNNs and MLPs have attained remarkable successes in broad vision tasks thanks to their effective information fusion in the global scope. Ho…

cs.CV2023

Vector-based Representation is the Key: A Study on Disentanglement and Compositional Generalization

Tao Yang, Yuwang Wang, Cuiling Lan +2

Recognizing elementary underlying concepts from observations (disentanglement) and generating novel combinations of these concepts (compositional generalization) are fundamental ab…

cs.CV20231 cited

Learning Distortion Invariant Representation for Image Restoration from A Causality Perspective

Xin Li, Bingchen Li, Xin Jin +2

In recent years, we have witnessed the great advancement of Deep neural networks (DNNs) in image restoration. However, a critical limitation is that they cannot generalize well to…

cs.LG20231 cited

Versatile Neural Processes for Learning Implicit Neural Representations

Zongyu Guo, Cuiling Lan, Zhizheng Zhang +2

Representing a signal as a continuous function parameterized by neural network (a.k.a. Implicit Neural Representations, INRs) has attracted increasing attention in recent years. Ne…