3 papers
cs.IR2025
Factorized Transport Alignment for Multimodal and Multiview E-commerce Representation Learning
Xiwen Chen, Yen-Chieh Lien, Susan Liu +4
The rapid growth of e-commerce requires robust multimodal representations that capture diverse signals from user-generated listings. Existing vision-language models (VLMs) typicall…
cs.CV2025
Fast 2DGS: Efficient Image Representation with Deep Gaussian Prior
Hao Wang, Ashish Bastola, Chaoyi Zhou +5
As generative models become increasingly capable of producing high-fidelity visual content, the demand for efficient, interpretable, and editable image representations has grown su…
cs.CV2025
Diffusion Prism: Enhancing Diversity and Morphology Consistency in Mask-to-Image Diffusion
Hao Wang, Xiwen Chen, Ashish Bastola +2
The emergence of generative AI and controllable diffusion has made image-to-image synthesis increasingly practical and efficient. However, when input images exhibit low entropy and…