2 papers
cs.LG2026
Directional Reasoning Trajectory Change (DRTC): Identifying Critical Trace Segments in Reasoning Models
Waldemar Chang
Understanding how language models carry out long-horizon reasoning remains an open challenge. Existing interpretability methods often highlight tokens correlated with an answer, bu…
cs.CL2025
Fusion Steering: Prompt-Specific Activation Control
Waldemar Chang, Alhassan Yasin
We present Fusion Steering, an activation steering methodology that improves factual accuracy in large language models (LLMs) for question-answering (QA) tasks. This approach intro…