1 paper · 1 filter
Oliver Bentham, Nathan Stringham, Ana Marasović
Understanding the extent to which Chain-of-Thought (CoT) generations align with a large language model's (LLM) internal computations is critical for deciding whether to trust an LL…