4 papers
AesBiasBench: Evaluating Bias and Alignment in Multimodal Language Models for Personalized Image Aesthetic Assessment
Kun Li, Lai-Man Po, Hongzheng Yang +3
Multimodal Large Language Models (MLLMs) are increasingly applied in Personalized Image Aesthetic Assessment (PIAA) as a scalable alternative to expert evaluations. However, their…
Contrastive Spatio-Temporal Pretext Learning for Self-supervised Video Representation
Yujia Zhang, Lai-Man Po, Xuyuan Xu +5
Spatio-temporal representation learning is critical for video self-supervised representation. Recent approaches mainly use contrastive learning and pretext tasks. However, these ap…
VCGAN: Video Colorization with Hybrid Generative Adversarial Network
Yuzhi Zhao, Lai-Man Po, Wing-Yin Yu +4
We propose a hybrid recurrent Video Colorization with Hybrid Generative Adversarial Network (VCGAN), an improved approach to video colorization using end-to-end learning. The VCGAN…
Spatial Content Alignment For Pose Transfer
Wing-Yin Yu, Lai-Man Po, Yuzhi Zhao +2
Due to unreliable geometric matching and content misalignment, most conventional pose transfer algorithms fail to generate fine-trained person images. In this paper, we propose a n…