2 papers
cs.CL2025
Embedding Domain Knowledge for Large Language Models via Reinforcement Learning from Augmented Generation
Chaojun Nie, Jun Zhou, Guanxiang Wang +2
Large language models (LLMs) often exhibit limited performance on domain-specific tasks due to the natural disproportionate representation of specialized information in their train…
cs.LG2023
KDSM: An uplift modeling framework based on knowledge distillation and sample matching
Chang Sun, Qianying Li, Guanxiang Wang +2
Uplift modeling aims to estimate the treatment effect on individuals, widely applied in the e-commerce platform to target persuadable customers and maximize the return of marketing…