4 papers · 1 filter
Exploring System 1 and 2 communication for latent reasoning in LLMs
Julian Coda-Forno, Zhuokai Zhao, Qiang Zhang +6
Should LLM reasoning live in a separate module, or within a single model's forward pass and representational space? We study dual-architecture latent reasoning, where a fluent Base…
Reference-Free Rating of LLM Responses via Latent Information
Leander Girrbach, Chi-Ping Su, Tankred Saanum +3
How reliable are single-response LLM-as-a-judge ratings without references, and can we obtain fine-grained, deterministic scores in this setting? We study the common practice of as…
A circuit for predicting hierarchical structure in-context in Large Language Models
Tankred Saanum, Can Demircan, Samuel J. Gershman +1
Large Language Models (LLMs) excel at in-context learning, the ability to use information provided as context to improve prediction of future tokens. Induction heads have been argu…
Automated scientific minimization of regret
Marcel Binz, Akshay K. Jagadish, Milena Rmus +1
We introduce automated scientific minimization of regret (ASMR) -- a framework for automated computational cognitive science. Building on the principles of scientific regret minimi…