3 papers
cs.CL2026
Improving LLMs via Validator-to-Generator Alignment
Juan Diego Rodriguez, Jocelyn Zhang, Katrin Erk +1
Large language models are inconsistent: varying prompts or including unrelated information can lead to unexpected changes in model outputs. The generator-validator (G-V) gap is one…
cs.CL2026
Scene Abstraction for Lexical Semantics: Structured Representations of Situated Meaning
Yejin Cho, Katrin Erk
Coffee and tea share many properties, yet they evoke strikingly different situations, atmospheres, and affective associations. These situated dimensions of word meaning are real an…
cs.CL2025
RankAlign: A Ranking View of the Generator-Validator Gap in Large Language Models
Juan Diego Rodriguez, Wenxuan Ding, Katrin Erk +1
Although large language models (LLMs) have become more capable and accurate across many tasks, some fundamental sources of unreliability remain in their behavior. One key limitatio…