2 papers
cs.LG2026
LoopFM: Learning frOm HistOrical RePresentations of Foundation Model for Recommendation
Shali Jiang, Hua Zheng, Boyang Liu +40
Knowledge distillation (KD) transfers a single scalar prediction from a large foundation model (FM) to compact vertical models (VMs), suffering from diminishing transfer ratio -- t…
cs.LG2026
Expressive Power of Implicit Models: Rich Equilibria and Test-Time Scaling
Jialin Liu, Lisang Ding, Stanley Osher +1
Implicit models, an emerging model class, compute outputs by iterating a single parameter block to a fixed point. This architecture realizes an infinite-depth, weight-tied network…