3 papers
econ.GN2026
Randomness in large language models: What researchers need to know (and report)
Guillaume Coqueret, Joan Llull, Florian Oswald +3
Large language models (LLMs) are increasingly used to generate data for research. Typical use cases are classifications, annotations, information extraction, and generation of nume…
cs.LG2025
The Challenger: When Do New Data Sources Justify Switching Machine Learning Models?
Vassilis Digalakis, Christophe Pérignon, Sébastien Saurin +1
Organizations often have an incumbent predictive model in production when new data sources become available. Because historical training data lack the new features, a challenger mo…
stat.ME2025
Global p-Values in Multi-Design Studies
Guillaume Coqueret, Yuming Zhang, Christophe Pérignon +2
Replicability issues -- referring to the difficulty or failure of independent researchers to corroborate the results of published studies -- have hindered the meaningful progressio…