1 paper
Gerrit Mutschlechner, Adam Jatowt
This study evaluates the forecasting performance of recent language models (LLMs) on binary forecasting questions. We first introduce a novel dataset of over 600 binary forecasting…