1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.LG2025★ 1 cited
RedStar: Does Scaling Long-CoT Data Unlock Better Slow-Reasoning Systems?
Haotian Xu, Xing Wu, Weinong Wang +11
Can scaling transform reasoning? In this work, we explore the untapped potential of scaling Long Chain-of-Thought (Long-CoT) data to 1000k samples, pioneering the development of a…
cs.CL2024
Untie the Knots: An Efficient Data Augmentation Strategy for Long-Context Pre-Training in Language Models
Junfeng Tian, Da Zheng, Yang Cheng +3
Large language models (LLM) have prioritized expanding the context window from which models can incorporate more information. However, training models to handle long contexts prese…