embeddings 1game theory 1language models 1large language models 1on-policy distillation 1rollout optimization 1strategic reasoning 1structured pruning 1text generation 1transfer learning 1
From the 2 of 2 linked papers with an AI index.
2 papers
cs.GT2026
Strategy, Not Payoffs: A Behavioural Embedding of Normal-Form Games
Joshua Caiata, Sreepriya Pulyassary, Xiang Li +1
The paper introduces a lightweight behavioural embedding for normal-form games, using Nash equilibrium entropy and response sensitivity, to predict how fine‑tuning large language m…
cs.LG2026
ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation
Qingyu Zhang, Qianhao Yuan, Hongyu Lin +7
The paper proposes ShortOPD, a short-to-long on-policy distillation method that recovers the generation quality of structured-pruned large language models by focusing training on e…