5 papers
Creative4U: MLLMs-based Advertising Creative Image Selector with Comparative Reasoning
Yukang Lin, Xiang Zhang, Shichang Jia +9
Creative image in advertising is the heart and soul of e-commerce platform. An eye-catching creative image can enhance the shopping experience for users, boosting income for advert…
MOON Embedding: Multimodal Representation Learning for E-commerce Search Advertising
Chenghan Fu, Daoze Zhang, Yukang Lin +8
We introduce MOON, our comprehensive set of sustainable iterative practices for multimodal representation learning for e-commerce applications. MOON has already been fully deployed…
QFFT, Question-Free Fine-Tuning for Adaptive Reasoning
Wanlong Liu, Junxiao Xu, Fei Yu +7
Recent advancements in Long Chain-of-Thought (CoT) reasoning models have improved performance on complex tasks, but they suffer from overthinking, which generates redundant reasoni…
NTIRE 2025 challenge on Text to Image Generation Model Quality Assessment
Shuhao Han, Haotian Fan, Fangyuan Kong +112
This paper reports on the NTIRE 2025 challenge on Text to Image (T2I) generation model quality assessment, which will be held in conjunction with the New Trends in Image Restoratio…
HVIS: A Human-like Vision and Inference System for Human Motion Prediction
Kedi Lyu, Haipeng Chen, Zhenguang Liu +3
Grasping the intricacies of human motion, which involve perceiving spatio-temporal dependence and multi-scale effects, is essential for predicting human motion. While humans inhere…