25 citations · 29 across the 6 of their papers we have counts for
4 papers · 1 filter
MACEval: A Multi-Agent Continual Evaluation Network for Large Models
Zijian Chen, Yuze Sun, Yuan Tian +2
Hundreds of benchmarks dedicated to evaluating large models have been presented over the past few years. However, most of them remain closed-ended and are prone to overfitting due…
Just Noticeable Difference for Large Multimodal Models
Zijian Chen, Yuan Tian, Yuze Sun +5
Just noticeable difference (JND), the minimum change that the human visual system (HVS) can perceive, has been studied for decades. Although recent work has extended this line of r…
OBI-Bench: Can LMMs Aid in Study of Ancient Script on Oracle Bones?
Zijian Chen, Tingzhu Chen, Wenjun Zhang +1
We introduce OBI-Bench, a holistic benchmark crafted to systematically evaluate large multi-modal models (LMMs) on whole-process oracle bone inscriptions (OBI) processing tasks dem…
Exploring Rich Subjective Quality Information for Image Quality Assessment in the Wild
Xiongkuo Min, Yixuan Gao, Yuqin Cao +4
Traditional in the wild image quality assessment (IQA) models are generally trained with the quality labels of mean opinion score (MOS), while missing the rich subjective quality i…