Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
The Illusion of Latent Generalization: Bi-directionality and the Reversal Curse
Julian Coda-Forno, Jane X. Wang, Arslan Chaudhry
The reversal curse describes a failure of autoregressive language models to retrieve a fact in reverse order (e.g., training on ``'' but failing on ``''). Recent work…
cs.CL2024
Machine Psychology
Thilo Hagendorff, Ishita Dasgupta, Marcel Binz +5
Large language models (LLMs) show increasingly advanced emergent capabilities and are being incorporated across various societal domains. Understanding their behavior and reasoning…