3 papers
cs.LG2026
What Makes a Terminal-Bench Task Hard? Separating Genuine Hardness from Fake-Hardness on an Adjudicated Agentic Corpus
Edward Lue Chee Lip, Boden Moraski, Tim Knappe +4
Frontier benchmarks need tasks that current models cannot solve. But a task that no model solves is not automatically a hard task. The same zero pass rate can come from a real capa…
cs.CL2026
On Mitigation of Subliminal Learning in Large Language Models
Atsushi Yanagisawa, Brendan Gho, Rajendran Ramesh Babu Manoj Narender +3
Knowledge distillation can transmit unintended behavioral traits from a teacher model to a student through training data that appear semantically unrelated to those traits, a pheno…
cs.LG2024
What is the Relationship between Tensor Factorizations and Circuits (and How Can We Exploit it)?
Lorenzo Loconte, Antonio Mari, Gennaro Gala +5
This paper establishes a rigorous connection between circuit representations and tensor factorizations, two seemingly distinct yet fundamentally related areas. By connecting these…