1 paper
Andrii Shportko, Shubham Bhokare, Ahmed Zeyad A Alzahrani +3
Fine-tuning through RL reshapes the internal representations of language models to enable agentic behaviors such as tool use, yet the mechanistic basis of these changes remains poo…