1 paper
Yindong Wang, Martin PreiÃ, Margarita Bugueño +4
The mechanisms underlying scientific confabulation in Large Language Models (LLMs) remain poorly understood. We introduce ReFACT (Reddit False And Correct Texts), a benchmark of 1,…