7 citations · 13 across the 3 of their papers we have counts for
3 papers
cs.CL2024★ 7 cited
Continual Pre-Training for Cross-Lingual LLM Adaptation: Enhancing Japanese Language Capabilities
Kazuki Fujii, Taishi Nakamura, Mengsay Loem +7
Cross-lingual continual pre-training of large language models (LLMs) initially trained on English corpus allows us to leverage the vast amount of English language resources and red…
cs.CL2024★ 5 cited
Building a Large Japanese Web Corpus for Large Language Models
Naoaki Okazaki, Kakeru Hattori, Hirai Shota +7
Open Japanese large language models (LLMs) have been trained on the Japanese portions of corpora such as CC-100, mC4, and OSCAR. However, these corpora were not created for the qua…
cs.CV2024★ 1 cited
Heron-Bench: A Benchmark for Evaluating Vision Language Models in Japanese
Yuichi Inoue, Kento Sasaki, Yuma Ochi +3
Vision Language Models (VLMs) have undergone a rapid evolution, giving rise to significant advancements in the realm of multimodal understanding tasks. However, the majority of the…