collaborators

6 papers

cs.MM2026

Subjective and Objective Quality-of-Experience Evaluation Study for Live Video Streaming

Zehao Zhu, Wei Sun, Jun Jia +7

In recent years, live video streaming has gained widespread popularity across various social media platforms. Quality of experience (QoE), which reflects end-users' satisfaction an…

cs.CV2026

Evaluating Image Editing with LLMs: A Comprehensive Benchmark and Intermediate-Layer Probing Approach

Shiqi Gao, Zitong Xu, Kang Fu +3

Evaluating text-guided image editing (TIE) methods remains a challenging problem, as reliable assessment should simultaneously consider perceptual quality, alignment with textual i…

cs.CV2026

Enhancing Image Quality Assessment Ability of LMMs via Retrieval-Augmented Generation

Kang Fu, Huiyu Duan, Zicheng Zhang +5

Large Multimodal Models (LMMs) have recently shown remarkable promise in low-level visual perception tasks, particularly in Image Quality Assessment (IQA), demonstrating strong zer…

cs.RO2025

Data Assessment for Embodied Intelligence

Jiahao Xiao, Bowen Yan, Jianbo Zhang +4

In embodied intelligence, datasets play a pivotal role, serving as both a knowledge repository and a conduit for information transfer. The two most critical attributes of a dataset…

cs.CL2025

Information Density Principle for MLLM Benchmarks

Chunyi Li, Xiaozhe Li, Zicheng Zhang +8

With the emergence of Multimodal Large Language Models (MLLMs), hundreds of benchmarks have been developed to ensure the reliability of MLLMs in downstream tasks. However, the eval…

cs.CV2025

Multi-Dimensional Quality Assessment for Text-to-3D Assets: Dataset and Model

Kang Fu, Huiyu Duan, Zicheng Zhang +4

Recent advancements in text-to-image (T2I) generation have spurred the development of text-to-3D asset (T23DA) generation, leveraging pretrained 2D text-to-image diffusion models f…