2 papers
cs.LG2025
AIRepr: An Analyst-Inspector Framework for Evaluating Reproducibility of LLMs in Data Science
Qiuhai Zeng, Claire Jin, Xinyue Wang +2
Large language models (LLMs) are increasingly used to automate data analysis through executable code generation. Yet, data science tasks often admit multiple statistically valid so…
cs.HC2024
Evaluating Fairness in Black-box Algorithmic Markets: A Case Study of Ride Sharing in Chicago
Yuhan Liu, Yuhan Zheng, Siyuan Zhang +1
This study examines fairness within the rideshare industry, focusing on both drivers' wages and riders' trip fares. Through quantitative analysis, we found that drivers' hourly wag…