2 papers
cs.CL2024
Towards Better Text-to-Image Generation Alignment via Attention Modulation
Yihang Wu, Xiao Cao, Kaixin Li +4
In text-to-image generation tasks, the advancements of diffusion models have facilitated the fidelity of generated results. However, these models encounter challenges when processi…
cs.CV2023
Class-level Structural Relation Modelling and Smoothing for Visual Representation Learning
Zitan Chen, Zhuang Qi, Xiao Cao +3
Representation learning for images has been advanced by recent progress in more complex neural models such as the Vision Transformers and new learning theories such as the structur…