From the 1 of 7 linked papers with an AI index.
7 papers
Reasoning Errors Have a Region and a Direction in the Residual-Stream Trajectory of LLMs
Hamed Damirchi, Ignacio Meza De la Jara, Damith Ranasinghe +2
As language models are increasingly used for tasks that require verifiable reasoning, reliably distinguishing sound reasoning from flawed reasoning has become an important practica…
Representation Trajectories Matters: Complementary Evidence for OOD Detection and Image Classification
Ignacio M. De la Jara, Cristian Rodriguez-Opazo, Hamed Damirchi +2
The paper investigates how the step‑by‑step changes in a vision model’s internal representations (representation trajectories) can be used to improve out‑of‑distribution detection…
Vertical Fusion: Condensing Internal Representations for Robust ViT Classification
Francesco Di Salvo, Shyam Nandan Rai, Hamed Damirchi +4
Despite exposing rich intermediate representations, Vision Transformers (ViTs) are almost exclusively utilized as black-box feature extractors, where only the last layer is conside…
Truth as a Trajectory: What Internal Representations Reveal About Large Language Model Reasoning
Hamed Damirchi, Ignacio Meza De la Jara, Ehsan Abbasnejad +3
Existing explainability methods for Large Language Models (LLMs) typically treat hidden states as static points in activation space, assuming that correct and incorrect inferences…
Decomposing Task Vectors for Refined Model Editing
Hamed Damirchi, Ehsan Abbasnejad, Zhen Zhang +1
Large pre-trained models have transformed machine learning, yet adapting these models effectively to exhibit precise, concept-specific behaviors remains a significant challenge. Ta…
The Quest for Winning Tickets in Low-Rank Adapters
Hamed Damirchi, Cristian Rodriguez-Opazo, Ehsan Abbasnejad +2
The Lottery Ticket Hypothesis (LTH) suggests that over-parameterized neural networks contain sparse subnetworks ("winning tickets") capable of matching full model performance when…