7 citations · 7 across the 2 of their papers we have counts for
2 papers
cs.AI2026
Evaluating and Understanding Scheming Propensity in LLM Agents
Mia Hopman, Jannes Elstner, Maria Avramidou +2
As frontier language models are increasingly deployed as autonomous agents pursuing complex, long-term objectives, there is increased risk of scheming: agents covertly pursuing mis…
hep-ph2022★ 7 cited
Picking the low-hanging fruit: testing new physics at scale with active learning
Juan Rocamonde, Louie Corpe, Gustavs Zilgalvis +2
Since the discovery of the Higgs boson, testing the many possible extensions to the Standard Model has become a key challenge in particle physics. This paper discusses a new method…