2 papers
cs.CL2026
Post-training makes large language models less human-like
Marcel Binz, Elif Akata, Abdullah Almaatouq +76
Large language models (LLMs) are increasingly used as surrogates for human participants, but it remains unclear which models best capture human behavior and why. To address this, w…
cs.NE2024
Evolving choice hysteresis in reinforcement learning: comparing the adaptive value of positivity bias and gradual perseveration
Isabelle Hoxha, Leo Sperber, Stefano Palminteri
The tendency of repeating past choices more often than expected from the history of outcomes has been repeatedly empirically observed in reinforcement learning experiments. It can…