collaborators

8 papers

cs.IR2025

Factorized Transport Alignment for Multimodal and Multiview E-commerce Representation Learning

Xiwen Chen, Yen-Chieh Lien, Susan Liu +4

The rapid growth of e-commerce requires robust multimodal representations that capture diverse signals from user-generated listings. Existing vision-language models (VLMs) typicall…

cs.CV2025

Fast 2DGS: Efficient Image Representation with Deep Gaussian Prior

Hao Wang, Ashish Bastola, Chaoyi Zhou +5

As generative models become increasingly capable of producing high-fidelity visual content, the demand for efficient, interpretable, and editable image representations has grown su…

cs.CV2025

Diffusion Prism: Enhancing Diversity and Morphology Consistency in Mask-to-Image Diffusion

Hao Wang, Xiwen Chen, Ashish Bastola +2

The emergence of generative AI and controllable diffusion has made image-to-image synthesis increasingly practical and efficient. However, when input images exhibit low entropy and…

cs.LG2024

Geographical Information Alignment Boosts Traffic Analysis via Transpose Cross-attention

Xiangyu Jiang, Xiwen Chen, Hao Wang +1

Traffic accident prediction is crucial for enhancing road safety and mitigating congestion, and recent Graph Neural Networks (GNNs) have shown promise in modeling the inherent grap…

cs.CV2024

Many-MobileNet: Multi-Model Augmentation for Robust Retinal Disease Classification

Hao Wang, Wenhui Zhu, Xuanzhao Dong +9

In this work, we propose Many-MobileNet, an efficient model fusion strategy for retinal disease classification using lightweight CNN architecture. Our method addresses key challeng…

cs.LG2024

Enhancing Graph Neural Networks in Large-scale Traffic Incident Analysis with Concurrency Hypothesis

Xiwen Chen, Sayed Pedram Haeri Boroujeni, Xin Shu +2

Despite recent progress in reducing road fatalities, the persistently high rate of traffic-related deaths highlights the necessity for improved safety interventions. Leveraging lar…