1 paper
Amelie Knecht, Lucas Florin, Thilo Hagendorff
Large reasoning models (LRMs) sometimes note in their chain of thought (CoT) that they may be under evaluation. Researchers worry that this verbalised evaluation awareness (VEA) ca…