3 papers
cs.LG2026
Speculative Decoding at Temperature Zero: A Scoped Safety-Invariance Screen with a 48,072-Sample Expansion
Sahil Kadadekar
Speculative decoding accelerates inference by letting a draft model propose tokens for a target model to verify, raising a concrete safety question: at temperature zero, can draft-…
cs.LG2026
Quality Is Not a Safety Proxy Under Quantization
Sahil Kadadekar
Quantized checkpoints are often screened first with quality metrics and only later, if at all, with direct safety tests. This paper audits that shortcut on a matched 51-row matrix…
cs.LG2026
A Paired Testing Protocol for Batch-Conditioned Refusal Robustness in LLM Serving
Sahil Kadadekar
Safety evaluations of language models often treat serving configuration as fixed background infrastructure, but batch condition is an untested treatment variable whenever the same…