2 papers
cs.CV2026
OmniFashion: Towards Generalist Fashion Intelligence via Multi-Task Vision-Language Learning
Zhengwei Yang, Andi Long, Hao Li +3
Fashion intelligence spans multiple tasks, i.e., retrieval, recommendation, recognition, and dialogue, yet remains hindered by fragmented supervision and incomplete fashion annotat…
cs.AI2025
VEGAS: Towards Visually Explainable and Grounded Artificial Social Intelligence
Hao Li, Hao Fei, Zechao Hu +2
Social Intelligence Queries (Social-IQ) serve as the primary multimodal benchmark for evaluating a model's social intelligence level. While impressive multiple-choice question(MCQ)…