1 paper
Dennis Wei, Yannis Belkhiter, Erik Miehling +1
Understanding the causal structure of a language model's thought process is a problem of significant importance for both transparency and safety. In this work, we take a local appr…