activity
20242026
collaborators

12 papers

cs.CR2026

Rendering on Real Silicon: GPU Render-Timing as a Passive, AI-Resistant CAPTCHA Signal

David Noever, Forrest McKee

Conventional CAPTCHAs pose puzzles that modern AI systems increasingly solve, while behavioral and cryptographic-attestation defenses carry privacy or enrollment costs. We investig…

cs.AI2025

Moravec's Paradox: Towards an Auditory Turing Test

David Noever, Forrest McKee

This research work demonstrates that current AI systems fail catastrophically on auditory tasks that humans perform effortlessly. Drawing inspiration from Moravec's paradox (i.e.,…

cs.CR2025

Favicon Trojans: Executable Steganography Via Ico Alpha Channel Exploitation

David Noever, Forrest McKee

This paper presents a novel method of executable steganography using the alpha transparency layer of ICO image files to embed and deliver self-decompressing JavaScript payloads wit…

cs.AI2025

Can AI Freelancers Compete? Benchmarking Earnings, Reliability, and Task Success at Scale

David Noever, Forrest McKee

This study explores Large Language Models (LLMs) as autonomous agents for real-world tasks, including freelance software development. This work presents a new benchmark that evalua…

cs.LG2025

Alpha Excel Benchmark

David Noever, Forrest McKee

This study presents a novel benchmark for evaluating Large Language Models (LLMs) using challenges derived from the Financial Modeling World Cup (FMWC) Excel competitions. We intro…

cs.CL2025

Forbidden Science: Dual-Use AI Challenge Benchmark and Scientific Refusal Tests

David Noever, Forrest McKee

The development of robust safety benchmarks for large language models requires open, reproducible datasets that can measure both appropriate refusal of harmful content and potentia…