2 papers
cs.CL2025
Do different prompting methods yield a common task representation in language models?
Guy Davidson, Todd M. Gureckis, Brenden M. Lake +1
Demonstrations and instructions are two primary approaches for prompting language models to perform in-context learning (ICL) tasks. Do identical tasks elicited in different ways r…
cs.AI2025
SAGE-Eval: Evaluating LLMs for Systematic Generalizations of Safety Facts
Chen Yueh-Han, Guy Davidson, Brenden M. Lake
Do LLMs robustly generalize critical safety facts to novel situations? Lacking this ability is dangerous when users ask naive questions. For instance, "I'm considering packing melo…