Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
ATLAS: Verifier-Guided Adaptive Latent Activation Steering for Efficient LLM Reasoning
Tuc Nguyen, Thai Le
Recent work on activation and latent steering has demonstrated that modifying internal representations can effectively guide large language models (LLMs) toward improved reasoning…
cs.LG2026
Beyond Linear Activation Steering: Invertible Latent Transformations for Controlling LLM Behavior
Tuc Nguyen, Thai Le
Activation steering provides a lightweight inference-time mechanism for controlling large language models (LLMs) by modifying their internal activation vectors toward desired behav…