Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
BODHI: Do LLMs Branch Out and Discover Heterogeneous Inferences?
Soumadeep Saha, Krish Sharma, Akshay Chaturvedi +1
Although reinforcement learning with verifiable rewards (RLVR) has improved the performance of large language models (LLMs) across a variety of reasoning tasks, there is significan…
cs.CL2024
On Explaining with Attention Matrices
Omar Naim, Nicholas Asher
This paper explores the much discussed, possible explanatory link between attention weights (AW) in transformer models and predicted output. Contrary to intuition and early researc…