2 papers
cs.AI2026
Epistemic Norms for AI Safety and Alignment Research
Keivan Navaie
Mainstream AI research emphasises capability growth and tolerates low failure rates when average-case performance is high. AI safety and alignment research has a different mission:…
cs.CR2026
Now You (Still) See Me: Detecting Evasive Steganographic Payloads in LLMs
Charles Westphal, Timothy Douglas, Keivan Navaie +2
Large language models can be fine-tuned to encode prompt-borne secrets into fluent, seemingly benign outputs. This creates a steganographic exfiltration risk that is difficult to d…