14 citations · 17 across the 5 of their papers we have counts for
3 papers · 1 filter
Connective Viewpoints of Signal-to-Noise Diffusion Models
Khanh Doan, Long Tung Vuong, Tuan Nguyen +5
Diffusion models (DM) have become fundamental components of generative models, excelling across various domains such as image creation, audio generation, and complex data interpola…
Vision Transformer Visualization: What Neurons Tell and How Neurons Behave?
Van-Anh Nguyen, Khanh Pham Dinh, Long Tung Vuong +4
Recently vision transformers (ViT) have been applied successfully for various tasks in computer vision. However, important questions such as why they work or how they behave still…
MoVQ: Modulating Quantized Vectors for High-Fidelity Image Generation
Chuanxia Zheng, Long Tung Vuong, Jianfei Cai +1
Although two-stage Vector Quantized (VQ) generative models allow for synthesizing high-fidelity and high-resolution images, their quantization operator encodes similar patches with…