Showing cs.CLShow all
3 papers · 1 filter
cs.CL2026
Self-Generated Text Recognition: Quality Heuristics, Cross-Task Transfer, and Downstream Bias in LLM Evaluation
Jesse St. Amand, Callum Canavan, Sohaib Imran +5
Self-Generated Text Recognition (SGTR)--the ability of an LLM to identify its own outputs--poses risks to AI safeguards that rely on LLMs as evaluators or monitors: an LLM may reco…
cs.CL2025
Out-of-Context Abduction: LLMs Make Inferences About Procedural Data Leveraging Declarative Facts in Earlier Training Data
Sohaib Imran, Rob Lamb, Peter M. Atkinson
Large language models (LLMs) are trained on large corpora, yet it is unclear whether they can reason about the information present within their training data. We design experiments…
cs.CL2025
Are LLM Belief Updates Consistent with Bayes' Theorem?
Sohaib Imran, Ihor Kendiukhov, Matthew Broerman +4
Do larger and more capable language models learn to update their "beliefs" about propositions more consistently with Bayes' theorem when presented with evidence in-context? To test…