2 papers
cs.AI2026
If It's Nice, Do It Twice: We Should Try Iterative Corpus Curation
Robin Young
Recent work demonstrates that filtering harmful content from pretraining data improves model safety without degrading capabilities. We propose a natural extension: do it again. A m…
cs.CL2026
NP-Hard Lower Bound Complexity for Semantic Self-Verification
Robin Young
We model Semantic Self-Verification (SSV) as the problem of determining whether a statement accurately characterizes its own semantic properties within a given interpretive framewo…