2 papers
cs.AI2026
Process Matters more than Output for Distinguishing Humans from Machines
Milena Rmus, Mathew D. Hardy, Thomas L. Griffiths +1
Reliable human-machine discrimination is becoming increasingly important as large language models and autonomous agents are deployed in online settings. Existing approaches evaluat…
cs.CL2024
When a language model is optimized for reasoning, does it still show embers of autoregression? An analysis of OpenAI o1
R. Thomas McCoy, Shunyu Yao, Dan Friedman +2
In "Embers of Autoregression" (McCoy et al., 2023), we showed that several large language models (LLMs) have some important limitations that are attributable to their origins in ne…