2 citations · 3 across the 2 of their papers we have counts for
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
SCALE: Upscaled Continual Learning of Large Language Models
Jin-woo Lee, Junhwa Choi, Bongkyu Hwang +8
We revisit continual pre-training for large language models and argue that progress now depends more on scaling the right structure than on scaling parameters alone. We introduce S…
cs.CL2023★ 2 cited
Shuffle & Divide: Contrastive Learning for Long Text
Joonseok Lee, Seongho Joe, Kyoungwon Park +4
We propose a self-supervised learning method for long text documents based on contrastive learning. A key to our method is Shuffle and Divide (SaD), a simple text augmentation algo…