5 citations · 5 across the 2 of their papers we have counts for
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
The Illusion of Latent Generalization: Bi-directionality and the Reversal Curse
Julian Coda-Forno, Jane X. Wang, Arslan Chaudhry
The reversal curse describes a failure of autoregressive language models to retrieve a fact in reverse order (e.g., training on ``'' but failing on ``''). Recent work…
cs.CL2024★ 5 cited
CogBench: a large language model walks into a psychology lab
Julian Coda-Forno, Marcel Binz, Jane X. Wang +1
Large language models (LLMs) have significantly advanced the field of artificial intelligence. Yet, evaluating them comprehensively remains challenging. We argue that this is partl…