2 papers
cs.CV2026
Agentic Retoucher for Text-To-Image Generation
Shaocheng Shen, Jianfeng Liang, Chunlei Cai +5
Text-to-image (T2I) diffusion models such as SDXL and FLUX have achieved impressive photorealism, yet small-scale distortions remain pervasive in limbs, face, text and so on. Exist…
cs.CV2025
SMC++: Masked Learning of Unsupervised Video Semantic Compression
Yuan Tian, Xiaoyue Ling, Cong Geng +3
Most video compression methods focus on human visual perception, neglecting semantic preservation. This leads to severe semantic loss during the compression, hampering downstream v…