35 citations · 47 across the 2 of their papers we have counts for
9 papers
E-MD3C: Taming Masked Diffusion Transformers for Efficient Zero-Shot Object Customization
Trung X. Pham, Zhang Kang, Ji Woo Hong +2
We propose E-MD3C (fficient asked iffusion Transformer with Disentangled onditions and ompact $\underline…
Human Aesthetic Preference-Based Large Text-to-Image Model Personalization: Kandinsky Generation as an Example
Aven-Le Zhou, Yu-Ao Wang, Wei Wu +1
With the advancement of neural generative capabilities, the art community has actively embraced GenAI (generative artificial intelligence) for creating painterly content. Large tex…
Learning from Multi-Perception Features for Real-Word Image Super-resolution
Axi Niu, Kang Zhang, Trung X. Pham +4
Currently, there are two popular approaches for addressing real-world image super-resolution problems: degradation-estimation-based and blind-based methods. However, degradation-es…
CDPMSR: Conditional Diffusion Probabilistic Models for Single Image Super-Resolution
Axi Niu, Kang Zhang, Trung X. Pham +4
Diffusion probabilistic models (DPM) have been widely adopted in image-to-image translation to generate high-quality images. Prior attempts at applying the DPM to image super-resol…
On the Pros and Cons of Momentum Encoder in Self-Supervised Visual Representation Learning
Trung Pham, Chaoning Zhang, Axi Niu +2
Exponential Moving Average (EMA or momentum) is widely used in modern self-supervised learning (SSL) approaches, such as MoCo, for enhancing performance. We demonstrate that such m…
Semi-Supervised Video Inpainting with Cycle Consistency Constraints
Zhiliang Wu, Hanyu Xuan, Changchang Sun +2
Deep learning-based video inpainting has yielded promising results and gained increasing attention from researchers. Generally, these methods usually assume that the corrupted regi…