46 citations · 78 across the 9 of their papers we have counts for
15 papers
SteerVTE: Seamless Video Text Editing with Style and Glyph Control
Kai Zeng, Moran Li, Zhengwei Wang +6
Visual text editing aims to precisely modify text in images and videos while preserving stylistic consistency and visual realism. Despite significant advances in the image domain,…
Exploring the Global-to-Local Attention Scheme in Graph Transformers: An Empirical Study
Gang Wu, Zhengwei Wang
Graph Transformers (GTs) show considerable potential in graph representation learning. The architecture of GTs typically integrates Graph Neural Networks (GNNs) with global attenti…
TEAM-Net: Multi-modal Learning for Video Action Recognition with Partial Decoding
Zhengwei Wang, Qi She, Aljosa Smolic
Most of existing video action recognition models ingest raw RGB frames. However, the raw video stream requires enormous storage and contains significant temporal redundancy. Video…
Generative adversarial networks in time series: A survey and taxonomy
Eoin Brophy, Zhengwei Wang, Qi She +1
Generative adversarial networks (GANs) studies have grown exponentially in the past few years. Their impact has been seen mainly in the computer vision field with realistic image a…
ACTION-Net: Multipath Excitation for Action Recognition
Zhengwei Wang, Qi She, Aljosa Smolic
Spatial-temporal, channel-wise, and motion patterns are three complementary and crucial types of information for video action recognition. Conventional 2D CNNs are computationally…
IROS 2019 Lifelong Robotic Vision Challenge -- Lifelong Object Recognition Report
Qi She, Fan Feng, Qi Liu +33
This report summarizes IROS 2019-Lifelong Robotic Vision Competition (Lifelong Object Recognition Challenge) with methods and results from the top finalists (out of over~…