3 papers
cs.AI2026
QMFOL: Benchmarking Large Language Model Reasoning via Quantifiable Monadic First-Order Logic Test Case Generation
Xinyi Zheng, Ling Shi, Tianlong Yu +3
Large Language Models (LLMs) have made significant progress in reasoning, particularly in deductive reasoning, which is crucial for high-stakes decision-making. As models improve,…
econ.GN2025
Testing for Spillovers in Resource Conservation: Evidence from a Natural Field Experiment
Lorenz Goette, Zhi Hao Lim
This paper studies whether behavioral interventions designed to promote resource conservation in one domain generate spillovers in another. Using a natural field experiment involvi…
cs.SE2025
Large Language Models are overconfident and amplify human bias
Fengfei Sun, Ningke Li, Kailong Wang +1
Large language models (LLMs) are revolutionizing every aspect of society. They are increasingly used in problem-solving tasks to substitute human assessment and reasoning. LLMs are…