3 papers
cs.CV2025
LongT2IBench: A Benchmark for Evaluating Long Text-to-Image Generation with Graph-structured Annotations
Zhichao Yang, Tianjiao Gu, Jianjie Wang +4
The increasing popularity of long Text-to-Image (T2I) generation has created an urgent need for automatic and interpretable models that can evaluate the image-text alignment in lon…
eess.IV2025
TuningIQA: Fine-Grained Blind Image Quality Assessment for Livestreaming Camera Tuning
Xiangfei Sheng, Zhichao Duan, Xiaofeng Pan +4
Livestreaming has become increasingly prevalent in modern visual communication, where automatic camera quality tuning is essential for delivering superior user Quality of Experienc…
cs.CV2025
Language-Guided Visual Perception Disentanglement for Image Quality Assessment and Conditional Image Generation
Zhichao Yang, Leida Li, Pengfei Chen +2
Contrastive vision-language models, such as CLIP, have demonstrated excellent zero-shot capability across semantic recognition tasks, mainly attributed to the training on a large-s…