2 papers
cs.CV2025
EVLF-FM: Explainable Vision Language Foundation Model for Medicine
Yang Bai, Haoran Cheng, Yang Zhou +40
Despite the promise of foundation models in medical AI, current systems remain limited - they are modality-specific and lack transparent reasoning processes, hindering clinical ado…
cs.CV2024
Searching Priors Makes Text-to-Video Synthesis Better
Haoran Cheng, Liang Peng, Linxuan Xia +5
Significant advancements in video diffusion models have brought substantial progress to the field of text-to-video (T2V) synthesis. However, existing T2V synthesis model struggle t…