Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
MTAVG-Bench 2.0: Diagnosing Failure Modes of Cinematic Expressiveness in Multi-Talker Audio-Video Generation
Haitian Li, Yanghao Zhou, Heyan Huang +15
In recent years, Multi-Talker Audio-Video Generation (MTAVG) models have shown promising performance on fundamental metrics such as lip-sync and audio-visual alignment. However, th…
cs.AI2026
DeepSurvey-Bench: Evaluating Academic Value of Automatically Generated Scientific Surveys
Guo-Biao Zhang, Xian-Ling Mao, Ding-Yuan Liu +4
The rapid development of automated survey generation technology has made it increasingly important to establish a comprehensive benchmark to evaluate the quality of generated surve…
cs.AI2025
T2I-Eval-R1: Reinforcement Learning-Driven Reasoning for Interpretable Text-to-Image Evaluation
Zi-Ao Ma, Tian Lan, Rong-Cheng Tu +5
The rapid progress in diffusion-based text-to-image (T2I) generation has created an urgent need for interpretable automatic evaluation methods that can assess the quality of genera…