1 paper · 2 filters
Oliver Bentham, Nathan Stringham, Ana Marasović
Understanding the extent to which Chain-of-Thought (CoT) generations align with a large language model's (LLM) internal computations is critical for deciding whether to trust an LL…