Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Sequential Data Poisoning in LLM Post-Training
Jack Sanderson, Yihan Wang, Xiaoqian Lu +2
LLM post-training proceeds through multiple stages, e.g., supervised fine-tuning (SFT) followed by reinforcement learning from human feedback (RLHF) or direct preference optimizati…
cs.LG2025
Rethinking LLM Advancement: Compute-Dependent and Independent Paths to Progress
Jack Sanderson, Teddy Foley, Spencer Guo +2
Regulatory efforts to govern large language model (LLM) development have predominantly focused on restricting access to high-performance computational resources. This study evaluat…