89 citations · 90 across the 3 of their papers we have counts for
1 paper · 1 filter
Preferred Elements, :, Kenshin Abe +18
We introduce PLaMo-100B, a large-scale language model designed for Japanese proficiency. The model was trained from scratch using 2 trillion tokens, with architecture such as QK No…