Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
Forbidden Science: Dual-Use AI Challenge Benchmark and Scientific Refusal Tests
David Noever, Forrest McKee
The development of robust safety benchmarks for large language models requires open, reproducible datasets that can measure both appropriate refusal of harmful content and potentia…
cs.CL2024
The Impossible Test: A 2024 Unsolvable Dataset and A Chance for an AGI Quiz
David Noever, Forrest McKee
This research introduces a novel evaluation framework designed to assess large language models' (LLMs) ability to acknowledge uncertainty on 675 fundamentally unsolvable problems.…