2 papers
cs.LG2025
Data Uniformity Improves Training Efficiency and More, with a Convergence Framework Beyond the NTK Regime
Yuqing Wang, Shangding Gu
Data selection plays a crucial role in data-driven decision-making, including in large language models (LLMs), and is typically task-dependent. Properties such as data quality and…
cs.LG2025
RLBenchNet: The Right Network for the Right Reinforcement Learning Task
Ivan Smirnov, Shangding Gu
Reinforcement learning (RL) has seen significant advancements through the application of various neural network architectures. In this study, we systematically investigate the perf…