2 papers
cs.AI2026
Reasoning-Preserving Fine-Tuning of Post-RL LLMs with Null-Basis LoRA
Wenzhi Fang, Nicholas Tzou, Lazar Valkov +1
Reinforcement learning (RL)-based post-training has become an effective approach for eliciting reasoning capabilities in large language models (LLMs). However, adapting post-RL mod…
cs.CL2023
STEER: Semantic Turn Extension-Expansion Recognition for Voice Assistants
Leon Liyang Zhang, Jiarui Lu, Joel Ruben Antony Moniz +5
In the context of a voice assistant system, steering refers to the phenomenon in which a user issues a follow-up command attempting to direct or clarify a previous turn. We propose…