1 paper
Adam Karvonen, Euan Ong, Subhash Kantamneni +1
Many areas of AI research, such as language model interpretability and chain of thought faithfulness, seek to explain model behaviors. But what constitutes a "good" explanation? In…