2 papers
cs.CV2025
GiVE: Guiding Visual Encoder to Perceive Overlooked Information
Junjie Li, Jianghong Ma, Xiaofeng Zhang +2
Multimodal Large Language Models have advanced AI in applications like text-to-video generation and visual question answering. These models rely on visual encoders to convert non-t…
cs.CV2025
BC-GAN: A Generative Adversarial Network for Synthesizing a Batch of Collocated Clothing
Dongliang Zhou, Haijun Zhang, Jianghong Ma +1
Collocated clothing synthesis using generative networks has become an emerging topic in the field of fashion intelligence, as it has significant potential economic value to increas…