4 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.AI2024
HumanVLM: Foundation for Human-Scene Vision-Language Model
Dawei Dai, Xu Long, Li Yutang +2
Human-scene vision-language tasks are increasingly prevalent in diverse social applications, yet recent advancements predominantly rely on models specifically tailored to individua…
cs.CV2024★ 4 cited
15M Multimodal Facial Image-Text Dataset
Dawei Dai, YuTang Li, YingGe Liu +3
Currently, image-text-driven multi-modal deep learning models have demonstrated their outstanding potential in many fields. In practice, tasks centered around facial images have br…