1 paper
Toshiki Katsube, Taiga Fukuhara, Kenichiro Ando +3
This work addresses the scarcity of high-quality, large-scale resources for Japanese Vision-and-Language (V&L) modeling. We present a scalable and reproducible pipeline that integr…