14 citations · 17 across the 2 of their papers we have counts for
2 papers
cs.CV2022★ 3 cited
Vision Transformer Visualization: What Neurons Tell and How Neurons Behave?
Van-Anh Nguyen, Khanh Pham Dinh, Long Tung Vuong +4
Recently vision transformers (ViT) have been applied successfully for various tasks in computer vision. However, important questions such as why they work or how they behave still…
cs.CV2022★ 14 cited
MoVQ: Modulating Quantized Vectors for High-Fidelity Image Generation
Chuanxia Zheng, Long Tung Vuong, Jianfei Cai +1
Although two-stage Vector Quantized (VQ) generative models allow for synthesizing high-fidelity and high-resolution images, their quantization operator encodes similar patches with…