5 citations · 10 across the 11 of their papers we have counts for
1 paper · 1 filter
Katie Matton, Robert Osazuwa Ness, John Guttag +1
Large language models (LLMs) are capable of generating plausible explanations of how they arrived at an answer to a question. However, these explanations can misrepresent the model…