3 papers
cs.IR2026
Bridging Short Videos and Live Streams: Reasoning-Guided Multimodal LLMs for Cross-Domain Representation Learning
Le Zhang, Xiaolan Zhu, Yuchen Wang +9
As live streaming services grow, many platforms offer short videos and live streams to meet diverse needs. Short videos carry substantial traffic and rich behavior signals, whereas…
cs.IR2025
Multimodal Recommendation via Self-Corrective Preference Alignmen
Yalong Guan, Xiang Chen, Mingyang Wang +7
With the rapid growth of live streaming platforms, personalized recommendation systems have become pivotal in improving user experience and driving platform revenue. The dynamic an…
cs.CV2024
Human-Guided Image Generation for Expanding Small-Scale Training Image Datasets
Changjian Chen, Fei Lv, Yalong Guan +4
The performance of computer vision models in certain real-world applications (e.g., rare wildlife observation) is limited by the small number of available images. Expanding dataset…