1 paper · 1 filter
Casey O. Barkan, Sid Black, Oliver Sourbut
We investigate whether large language models (LLMs) can predict whether they will succeed on a given task and whether their predictions improve as they progress through multi-step…