Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
Multimodal LLMs under Pairwise Modalities
Yan Li, Yunlong Deng, Yuewen Sun +3
Despite the impressive results achieved by multimodal large language models (MLLMs), their training typically relies on jointly curated multimodal data, requiring substantial human…
cs.CV2025
Towards Self-Refinement of Vision-Language Models with Triangular Consistency
Yunlong Deng, Guangyi Chen, Tianpei Gu +4
Vision-Language Models (VLMs) integrate visual knowledge with the analytical capabilities of Large Language Models (LLMs) through supervised visual instruction tuning, using image-…
cs.CV2025
ProtoGS: Efficient and High-Quality Rendering with 3D Gaussian Prototypes
Zhengqing Gao, Dongting Hu, Jia-Wang Bian +5
3D Gaussian Splatting (3DGS) has made significant strides in novel view synthesis but is limited by the substantial number of Gaussian primitives required, posing challenges for de…