4 papers
Text-Guided Visual Representation Learning for Robust Multimodal E-Commerce Recommendation
Yufei Guo, Jing Ma, Tianlu Zhang +5
Multimodal item embeddings are crucial for e-commerce item-to-item (I2I) retrieval, yet real-world product images often contain promotional overlays and background clutter that inj…
OxygenREC: An Instruction-Following Generative Framework for E-commerce Recommendation
Xuegang Hao, Ming Zhang, Alex Li +30
Traditional recommendation systems suffer from inconsistency in multi-stage optimization objectives. Generative Recommendation (GR) mitigates them through an end-to-end framework;…
IC-Portrait: In-Context Matching for View-Consistent Personalized Portrait
Han Yang, Enis Simsar, Sotiris Anagnostidis +3
Existing diffusion models show great potential for identity-preserving generation. However, personalized portrait generation remains challenging due to the diversity in user profil…
High-Fidelity Virtual Try-on with Large-Scale Unpaired Learning
Han Yang, Yanlong Zang, Ziwei Liu
Virtual try-on (VTON) transfers a target clothing image to a reference person, where clothing fidelity is a key requirement for downstream e-commerce applications. However, existin…