6 citations · 6 across the 2 of their papers we have counts for
2 papers
cs.CL2026
Hidden Measurement Error in LLM Pipelines Distorts Annotation, Evaluation, and Benchmarking
Solomon Messing
LLM evaluations drive which models get deployed, what safety standards get adopted, which research conclusions get published, and how projections of AI's labor-market impact get ma…
cs.CY2024★ 6 cited
Web Scraping for Research: Legal, Ethical, Institutional, and Scientific Considerations
Megan A. Brown, Andrew Gruen, Gabe Maldoff +3
Scientists across disciplines often use data from the internet to conduct research, generating valuable insights about human behavior. However, as generative AI relying on massive…