2 papers
cs.CV2026
MaskAttn-SDXL: Controllable Region-Level Text-To-Image Generation
Yu Chang, Jiahao Chen, Anzhe Cheng +1
Diffusion models have achieved strong results in text-to-image generation, but important limitations remain as prompts become more structured and multi-object. On the architecture…
cs.CV2025
MaskAttn-UNet: A Mask Attention-Driven Framework for Universal Low-Resolution Image Segmentation
Anzhe Cheng, Chenzhong Yin, Yu Chang +4
Low-resolution image segmentation is crucial in real-world applications such as robotics, augmented reality, and large-scale scene understanding, where high-resolution data is ofte…