4 citations · 5 across the 3 of their papers we have counts for
3 papers
cs.AI2024
HumanVLM: Foundation for Human-Scene Vision-Language Model
Dawei Dai, Xu Long, Li Yutang +2
Human-scene vision-language tasks are increasingly prevalent in diverse social applications, yet recent advancements predominantly rely on models specifically tailored to individua…
cs.AI2024★ 1 cited
PA-LLaVA: A Large Language-Vision Assistant for Human Pathology Image Understanding
Dawei Dai, Yuanhui Zhang, Long Xu +4
The previous advancements in pathology image understanding primarily involved developing models tailored to specific tasks. Recent studies has demonstrated that the large vision-la…
cs.CV2024★ 4 cited
15M Multimodal Facial Image-Text Dataset
Dawei Dai, YuTang Li, YingGe Liu +3
Currently, image-text-driven multi-modal deep learning models have demonstrated their outstanding potential in many fields. In practice, tasks centered around facial images have br…