most citedQuadTree Attention for Vision Transformers

70 citations · 79 across the 6 of their papers we have counts for

collaborators

6 papers

cs.CV20223 cited

Latent Multi-Relation Reasoning for GAN-Prior based Image Super-Resolution

Jiahui Zhang, Fangneng Zhan, Yingchen Yu +3

Recently, single image super-resolution (SR) under large scaling factors has witnessed impressive progress by introducing pre-trained generative adversarial networks (GANs) as prio…

cs.CV20226 cited

RenderNet: Visual Relocalization Using Virtual Viewpoints in Large-Scale Indoor Environments

Jiahui Zhang, Shitao Tang, Kejie Qiu +6

Visual relocalization has been a widely discussed problem in 3D vision: given a pre-constructed 3D visual map, the 6 DoF (Degrees-of-Freedom) pose of a query image is estimated. Re…

cs.CV2022

Auto-regressive Image Synthesis with Integrated Quantization

Fangneng Zhan, Yingchen Yu, Rongliang Wu +4

Deep generative models have achieved conspicuous progress in realistic image synthesis with multifarious conditional inputs, while generating diverse yet high-fidelity images remai…

cs.CV2022

Towards Counterfactual Image Manipulation via CLIP

Yingchen Yu, Fangneng Zhan, Rongliang Wu +6

Leveraging StyleGAN's expressivity and its disentangled latent codes, existing methods can achieve realistic editing of different visual attributes such as age and gender of facial…

cs.CV202270 cited

QuadTree Attention for Vision Transformers

Shitao Tang, Jiahui Zhang, Siyu Zhu +1

Transformers have been successful in many vision tasks, thanks to their capability of capturing long-range dependency. However, their quadratic computational complexity poses a maj…

cs.DB2017

hMDAP: A Hybrid Framework for Multi-paradigm Data Analytical Processing on Spark

Xiaowang Zhang, Jiahui Zhang, Zhiyong Feng

We propose hMDAP, a hybrid framework for large-scale data analytical processing on Spark, to support multi-paradigm process (incl. OLAP, machine learning, and graph analysis etc.)…