2 papers
cs.CL2025
On the generalization of language models from in-context learning and finetuning: a controlled study
Andrew K. Lampinen, Arslan Chaudhry, Stephanie C. Y. Chan +7
Large language models exhibit exciting capabilities, yet can show surprisingly narrow generalization from finetuning. E.g. they can fail to generalize to simple reversals of relati…
cs.AI2024
CI-Bench: Benchmarking Contextual Integrity of AI Assistants on Synthetic Data
Zhao Cheng, Diane Wan, Matthew Abueg +6
Advances in generative AI point towards a new era of personalized applications that perform diverse tasks on behalf of users. While general AI assistants have yet to fully emerge,…