3 papers
cs.LG2025
Scalable Data Attribution via Forward-Only Test-Time Inference
Sibo Ma, Julian Nyarko
Data attribution seeks to trace model behavior back to the training examples that shaped it, enabling debugging, auditing, and data valuation at scale. Classical influence-function…
cs.CL2025
Identifying Emerging Concepts in Large Corpora
Sibo Ma, Julian Nyarko
We introduce a new method to identify emerging concepts in large text corpora. By analyzing changes in the heatmaps of the underlying embedding space, we are able to detect these c…
cs.CL2025
Breaking Down Bias: On The Limits of Generalizable Pruning Strategies
Sibo Ma, Alejandro Salinas, Peter Henderson +1
We employ model pruning to examine how LLMs conceptualize racial biases, and whether a generalizable mitigation strategy for such biases appears feasible. Our analysis yields sever…