1 paper
Haiwen Diao, Mingxuan Li, Silei Wu +6
The edifice of native Vision-Language Models (VLMs) has emerged as a rising contender to typical modular VLMs, shaped by evolving model architectures and training paradigms. Yet, t…