3 papers
cs.IR2026
Text-Guided Visual Representation Learning for Robust Multimodal E-Commerce Recommendation
Yufei Guo, Jing Ma, Tianlu Zhang +5
Multimodal item embeddings are crucial for e-commerce item-to-item (I2I) retrieval, yet real-world product images often contain promotional overlays and background clutter that inj…
cs.IR2025
OxygenREC: An Instruction-Following Generative Framework for E-commerce Recommendation
Xuegang Hao, Ming Zhang, Alex Li +30
Traditional recommendation systems suffer from inconsistency in multi-stage optimization objectives. Generative Recommendation (GR) mitigates them through an end-to-end framework;…
cs.CV2025
IC-Portrait: In-Context Matching for View-Consistent Personalized Portrait
Han Yang, Enis Simsar, Sotiris Anagnostidis +3
Existing diffusion models show great potential for identity-preserving generation. However, personalized portrait generation remains challenging due to the diversity in user profil…