1 paper
Jaemin Son, Sujin Choi, Inyong Yun
Recent progress in vision-language models (VLMs) has led to impressive results in document understanding tasks, but their high computational demands remain a challenge. To mitigate…