2 papers
cs.LG2026
SkillTFM: Gated Skill Evolution for Training-Free Adaptation of Tabular Foundation Models
Yi He, Zhengkang Guan, Anpeng Wu +3
Tabular data are ubiquitous in real-world applications and are crucial for data-driven prediction and decision-making across science, industry, finance, healthcare, and public serv…
cs.CL2026
Self-Distillation Zero: Self-Revision Turns Binary Rewards into Dense Supervision
Yinghui He, Simran Kaur, Adithya Bhaskar +7
Current post-training methods in verifiable settings fall into two categories. Reinforcement learning (RLVR) relies on binary rewards, which are broadly applicable and powerful, bu…