2 papers
cs.LG2026
Weightless Fine-Tuning: Personalizing LLMs via Logit-Space Transport
Bohan Zhang, Anqi Ni, Yixin Wang +1
Supervised fine-tuning (SFT) is a standard approach for adapting LLMs to a target distribution, but in settings such as personalization, where each author requires separate weight…
cs.CL2025
Policy Learning with a Natural Language Action Space: A Causal Approach
Bohan Zhang, Yixin Wang, Paramveer S. Dhillon
This paper introduces a novel causal framework for multi-stage decision-making in natural language action spaces where outcomes are only observed after a sequence of actions. While…