1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.AI2025
Back to the Baseline: Examining Baseline Effects on Explainability Metrics
Agustin Martin Picard, Thibaut Boissin, Varshini Subhash +2
Attribution methods are among the most prevalent techniques in Explainable Artificial Intelligence (XAI) and are usually evaluated and compared using Fidelity metrics, with Inserti…
cs.LG2023★ 1 cited
Why do universal adversarial attacks work on large language models?: Geometry might be the answer
Varshini Subhash, Anna Bialas, Weiwei Pan +1
Transformer based large language models with emergent capabilities are becoming increasingly ubiquitous in society. However, the task of understanding and interpreting their intern…