3 papers
stat.ME2026
Cluster-Robust Prediction-Powered Inference
David Broska, Michael Howes
Data collection is often costly or logistically demanding, limiting both the questions researchers can pursue and how precisely they can answer them. Prediction-powered inference (…
cs.CL2026
Post-training makes large language models less human-like
Marcel Binz, Elif Akata, Abdullah Almaatouq +76
Large language models (LLMs) are increasingly used as surrogates for human participants, but it remains unclear which models best capture human behavior and why. To address this, w…
cs.AI2026
This human study did not involve human subjects: Validating LLM simulations as behavioral evidence
Jessica Hullman, David Broska, Huaman Sun +1
A growing literature uses large language models (LLMs) as synthetic participants to generate cost-effective and nearly instantaneous responses in social science experiments. Howeve…