activity
20192026
most citedPerceptual Quality Assessment of Omnidirectional Images

159 citations · 290 across the 75 of their papers we have counts for

collaborators
Showing cs.CVShow all

72 papers · 1 filter

cs.CV2026

CamWorldQA: Perceptual Quality Assessment of Camera-Controlled World Video Generation

Yunhe Li, Likun Wu, Sijing Wu +5

Recent advances in generative video models have enabled camera-controlled world video generation, allowing models to synthesize videos under user-defined camera trajectories. Howev…

cs.CV2026

FMReward: Aligning and Evaluating Audio-Driven 3D Facial Animation with Human Preferences

Sijing Wu, Yunhao Li, Zhilin Gao +4

Audio-driven 3D facial animation is essential for advancing immersion and interactivity in virtual experiences. Although recent advances have shown promising capabilities, the trai…

cs.CV2026

Emo-Bench: A Scalable Benchmark for Multimodal Evoked and Expressed Emotion Understanding via Bayesian Pairwise Alignment

Lancheng Gao, Ziheng Jia, Shengyan Li +4

Understanding both expressed and evoked emotions is critical for multimodal large language models (MLLMs) to achieve comprehensive affect-aware interactions. However, existing benc…

cs.CV2026

MIEScore: Human-Aligned Evaluation for Multi-Source Image Editing

Zitong Xu, Huiyu Duan, Xinyun Zhang +7

Recent advances in unified multimodal models have significantly improved text-guided image editing abilities. In particular, models such as Nano-Banana-Pro and GPT-Image-2 demonstr…

cs.CV2026

Multi-Dimensional Quality Assessment for AI-Generated Human-Centric Videos: Dataset and Model

Sijing Wu, Yunhao Li, Huiyu Duan +4

AI-generated human-centric videos play a crucial role in a wide range of modern applications. However, they often suffer from quality issues and semantic mismatches, underscoring t…

cs.CV2026

LL-Bench: Rethinking Low-Level Vision Evaluation in the Era of Large-Scale Generative Models

Lu Liu, Huiyu Duan, Chenxin Zhu +6

Large-scale generative models have demonstrated remarkable capabilities across image generation and editing tasks. However, their performance in low-level vision tasks, which requi…