Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
TRACE: Trajectory-Based Safety Patch Learning for LLM Post-Training Realignment
Changyue Li, Jiaming He, Youliang Yuan +4
Fine-Tuning-as-a-Service (FTaaS) platforms let users train large language models (LLMs) on customized tasks, but this pipeline could erode models' safety alignment. In practice, se…
cs.LG2026
UTOPIA: Unlearnable Tabular Data via Decoupled Shortcut Embedding
Jiaming He, Fuming Luo, Hongwei Li +5
Unlearnable examples (UE) have emerged as a practical mechanism to prevent unauthorized model training on private vision data, while extending this protection to tabular data is no…