most citedCross-Scale Cost Aggregation for Stereo Matching

35 citations · 47 across the 2 of their papers we have counts for

collaborators

9 papers

cs.CV2025

E-MD3C: Taming Masked Diffusion Transformers for Efficient Zero-Shot Object Customization

Trung X. Pham, Zhang Kang, Ji Woo Hong +2

We propose E-MD3C (fficient asked iffusion Transformer with Disentangled onditions and ompact $\underline…

cs.AI2024

Human Aesthetic Preference-Based Large Text-to-Image Model Personalization: Kandinsky Generation as an Example

Aven-Le Zhou, Yu-Ao Wang, Wei Wu +1

With the advancement of neural generative capabilities, the art community has actively embraced GenAI (generative artificial intelligence) for creating painterly content. Large tex…

cs.CV20231 cited

Learning from Multi-Perception Features for Real-Word Image Super-resolution

Axi Niu, Kang Zhang, Trung X. Pham +4

Currently, there are two popular approaches for addressing real-world image super-resolution problems: degradation-estimation-based and blind-based methods. However, degradation-es…

eess.IV20232 cited

CDPMSR: Conditional Diffusion Probabilistic Models for Single Image Super-Resolution

Axi Niu, Kang Zhang, Trung X. Pham +4

Diffusion probabilistic models (DPM) have been widely adopted in image-to-image translation to generate high-quality images. Prior attempts at applying the DPM to image super-resol…

cs.CV20225 cited

On the Pros and Cons of Momentum Encoder in Self-Supervised Visual Representation Learning

Trung Pham, Chaoning Zhang, Axi Niu +2

Exponential Moving Average (EMA or momentum) is widely used in modern self-supervised learning (SSL) approaches, such as MoCo, for enhancing performance. We demonstrate that such m…

cs.CV2022

Semi-Supervised Video Inpainting with Cycle Consistency Constraints

Zhiliang Wu, Hanyu Xuan, Changchang Sun +2

Deep learning-based video inpainting has yielded promising results and gained increasing attention from researchers. Generally, these methods usually assume that the corrupted regi…