1 paper · 1 filter
Michael Lan, Philip Torr, Fazl Barez
While transformer models exhibit strong capabilities on linguistic tasks, their complex architectures make them difficult to interpret. Recent work has aimed to reverse engineer tr…