From the 1 of 7 linked papers with an AI index.
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Step-wise Adaptive Integration of Supervised Fine-tuning and Reinforcement Learning for Task-Specific LLMs
Jack Chen, Fazhong Liu, Naruto Liu +7
Large language models (LLMs) excel at mathematical reasoning and logical problem-solving. The current popular training paradigms primarily use supervised fine-tuning (SFT) and rein…
cs.LG2025
Model Inversion in Split Learning for Personalized LLMs: New Insights from Information Bottleneck Theory
Yunmeng Shu, Shaofeng Li, Tian Dong +2
Personalized Large Language Models (LLMs) have become increasingly prevalent, showcasing the impressive capabilities of models like GPT-4. This trend has also catalyzed extensive r…