3 papers
cs.LG2026
Soft Contamination Means Benchmarks Test Shallow Generalization
Ari Spiesberger, Juan J. Vazquez, Nicky Pochinkov +4
If LLM training data is polluted with benchmark test data, then benchmark performance gives biased estimates of out-of-distribution (OOD) generalization. Typical decontamination fi…
cs.CL2024
AI-AI Bias: large language models favor communications generated by large language models
Walter Laurito, Benjamin Davis, Peli Grietzer +3
Are large language models (LLMs) biased in favor of communications produced by LLMs, leading to possible antihuman discrimination? Using a classical experimental design inspired by…
cs.AI2023
Understanding and Controlling a Maze-Solving Policy Network
Ulisse Mini, Peli Grietzer, Mrinank Sharma +3
To understand the goals and goal representations of AI systems, we carefully study a pretrained reinforcement learning policy that solves mazes by navigating to a range of target s…