Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Multimodal LLMs under Pairwise Modalities
Yan Li, Yunlong Deng, Yuewen Sun +3
Despite the impressive results achieved by multimodal large language models (MLLMs), their training typically relies on jointly curated multimodal data, requiring substantial human…
cs.CV2025
Towards Self-Refinement of Vision-Language Models with Triangular Consistency
Yunlong Deng, Guangyi Chen, Tianpei Gu +4
Vision-Language Models (VLMs) integrate visual knowledge with the analytical capabilities of Large Language Models (LLMs) through supervised visual instruction tuning, using image-…