1 paper · 1 filter
William Walden, Miriam Wanner
Hint-based faithfulness evaluations have established that Large Reasoning Models (LRMs) may not say what they think: they do not always volunteer information about how key parts of…