2 papers
stat.ME2026
An Experimental Design Approach to Evaluating Agentic AI's Autonomous Model Discovery
Hao He, Xueying Liu, Chris J. Kuhlman +1
Large language model coding agents increasingly perform open-ended data modeling and analysis. These agents are stochastic and adaptive, and therefore their autonomous model discov…
stat.AP2025
StatLLM: A Dataset for Evaluating the Performance of Large Language Models in Statistical Analysis
Xinyi Song, Lina Lee, Kexin Xie +3
The coding capabilities of large language models (LLMs) have opened up new opportunities for automatic statistical analysis in machine learning and data science. However, before th…