4 papers
Child-Oriented AIGC Video Risk Reviewing: A Benchmark and Knowledge-Supported Iterative Reasoning Framework
Lewen Mi, Manyi Li, Yuling Sun +5
The rapid growth of Artificial Intelligence-generated content (AIGC) is reshaping video production and circulation, exposing children to an increasing volume of AIGC videos. Unlike…
Unifying Visual and Semantic Feature Spaces with Diffusion Models for Enhanced Cross-Modal Alignment
Yuze Zheng, Zixuan Li, Xiangxian Li +4
Image classification models often demonstrate unstable performance in real-world applications due to variations in image information, driven by differing visual perspectives of sub…
Class-level Structural Relation Modelling and Smoothing for Visual Representation Learning
Zitan Chen, Zhuang Qi, Xiao Cao +3
Representation learning for images has been advanced by recent progress in more complex neural models such as the Vision Transformers and new learning theories such as the structur…
Cross-Modal Content Inference and Feature Enrichment for Cold-Start Recommendation
Haokai Ma, Zhuang Qi, Xinxin Dong +3
Multimedia recommendation aims to fuse the multi-modal information of items for feature enrichment to improve the recommendation performance. However, existing methods typically in…