2 citations · 2 across the 5 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
CAREBench: A Child-Safety Risk Benchmark for Language Models
Kaavya Krishna-Kumar, Elaine Lau, Vaughn Robinson +6
How can we evaluate whether frontier AI systems recognize child-safety risks before they escalate into explicit harm? Existing child safety evaluations focus on child sexual abuse…
cs.LG2024
The Pragmatic Frames of Spurious Correlations in Machine Learning: Interpreting How and Why They Matter
Samuel J. Bell, Skyler Wang
Learning correlations from data forms the foundation of today's machine learning (ML) and artificial intelligence research. While contemporary methods enable the automatic discover…