3 papers
cs.CV2025
VARCO-VISION-2.0 Technical Report
Young-rok Cha, Jeongho Ju, SunYoung Park +3
We introduce VARCO-VISION-2.0, an open-weight bilingual vision-language model (VLM) for Korean and English with improved capabilities compared to the previous model VARCO-VISION-14…
cs.CV2025
Direction-Aware Diagonal Autoregressive Image Generation
Yijia Xu, Jianzhong Ju, Jian Luan +1
The raster-ordered image token sequence exhibits a significant Euclidean distance between index-adjacent tokens at line breaks, making it unsuitable for autoregressive generation.…
cs.CV2024
VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models
Jeongho Ju, Daeyoung Kim, SunYoung Park +1
In this paper, we introduce an open-source Korean-English vision-language model (VLM), VARCO-VISION. We incorporate a step-by-step training strategy that allows a model learn both…