3 papers
cs.CV2026
WeMM-Embedding: WeChat Multi-Modal Embedding Technical Report
Junjie Zhou, Ke Mei, Lei Li +3
Universal multimodal embeddings are becoming a core component of modern AI systems, enabling heterogeneous content to be represented in a shared space for applications such as retr…
cs.CV2024
Advancing Video Quality Assessment for AIGC
Xinli Yue, Jianhui Sun, Han Kong +9
In recent years, AI generative models have made remarkable progress across various domains, including text generation, image generation, and video generation. However, assessing th…
cs.CV2024
Revisiting Video Quality Assessment from the Perspective of Generalization
Xinli Yue, Jianhui Sun, Liangchao Yao +8
The increasing popularity of short video platforms such as YouTube Shorts, TikTok, and Kwai has led to a surge in User-Generated Content (UGC), which presents significant challenge…