activity
20242026
most citedLMME3DHF: Benchmarking and Evaluating Multimodal 3D Human Face Generation with LMMs

7 citations · 18 across the 50 of their papers we have counts for

collaborators
Showing cs.CVShow all

49 papers · 1 filter

cs.CV2026

MCIQA-2K: A Multi-Dimensional Dataset and No-Reference Quality Assessment Benchmark for Colorized Images

Yunkai Zhuang, Qihang Yan, Zicheng Zhang +1

Image colorization is an inherently ill-posed task, since a single grayscale image may correspond to multiple plausible colorized results. Consequently, conventional full-reference…

cs.CV2026

CamWorldQA: Perceptual Quality Assessment of Camera-Controlled World Video Generation

Yunhe Li, Likun Wu, Sijing Wu +5

Recent advances in generative video models have enabled camera-controlled world video generation, allowing models to synthesize videos under user-defined camera trajectories. Howev…

cs.CV2026

PCQA-R1: Advancing Generalized 3D Point Cloud Quality Assessment with Reinforcement Learning

Kangning Ye, Yunhao Li, Sijing Wu +2

No-reference point cloud quality assessment (PCQA) has been an active topic in recent years and is used to measure and optimize the visual experience of point clouds. However, larg…

cs.CV2026

FMReward: Aligning and Evaluating Audio-Driven 3D Facial Animation with Human Preferences

Sijing Wu, Yunhao Li, Zhilin Gao +4

Audio-driven 3D facial animation is essential for advancing immersion and interactivity in virtual experiences. Although recent advances have shown promising capabilities, the trai…

cs.CV2026

LLaVA-Assessor: Building the Foundation LMM For Visual Quality Assessment

Ziheng Jia, Zicheng Zhang, Jiaying Qian +2

Aligning with the human visual system~(HVS) in perceiving and evaluating the quality of visual signals is a central objective of machine-vision-based visual quality assessment syst…

cs.CV2026

3DGSI-Assessor: A Large-Scale Dataset and An LMM-based Method for 3D Gaussian Splatting Image Quality Assessment

Yuke Xing, Jiarui Wang, William Gordon +3

3D Gaussian Splatting (3DGS) has become a dominant representation for real-time novel view synthesis (NVS), yet its storage footprint makes compression indispensable for practical…