3 papers
cs.MM2025
XGC-AVis: Towards Audio-Visual Content Understanding with a Multi-Agent Collaborative System
Yuqin Cao, Xiongkuo Min, Yixuan Gao +4
In this paper, we propose XGC-AVis, a multi-agent framework that enhances the audio-video temporal alignment capabilities of multimodal large models (MLLMs) and improves the effici…
cs.CV2024
AIS 2024 Challenge on Video Quality Assessment of User-Generated Content: Methods and Results
Marcos V. Conde, Saman Zadtootaghaj, Nabajeet Barman +33
This paper reviews the AIS 2024 Video Quality Assessment (VQA) Challenge, focused on User-Generated Content (UGC). The aim of this challenge is to gather deep learning-based method…
cs.CV2023
Geometry-Aware Video Quality Assessment for Dynamic Digital Human
Zicheng Zhang, Yingjie Zhou, Wei Sun +2
Dynamic Digital Humans (DDHs) are 3D digital models that are animated using predefined motions and are inevitably bothered by noise/shift during the generation process and compress…