in-context learning 1language model evaluation 1prompt engineering 1self-consistency 1statistical consistency 1
From the 1 of 2 linked papers with an AI index.
2 papers
cs.CL2026
Partition, Prompt, Aggregate: Statistical Self-Consistency in Language Models
Patrik Wolf, Thomas Kleine Buening, Andreas Krause +1
The paper investigates whether large language models give probability estimates that satisfy the law of total probability when prompted for subpopulations, revealing frequent viola…
cs.LG2026
Specialization after Generalization: Towards Understanding Test-Time Training in Foundation Models
Jonas Hübotter, Patrik Wolf, Alexander Shevchenko +3
Recent empirical studies have explored the idea of continuing to train a model at test-time for a given task, known as test-time training (TTT), and have found it to yield signific…