3 papers
cs.CR2026
Building Better Environments for Autonomous Cyber Defence
Chris Hicks, Elizabeth Bates, Shae McFadden +12
In November 2025, the authors ran a workshop on the topic of what makes a good reinforcement learning (RL) environment for autonomous cyber defence (ACD). This paper details the kn…
cs.LG2025
Poisoning Attacks on LLMs Require a Near-constant Number of Poison Samples
Alexandra Souly, Javier Rando, Ed Chapman +10
Poisoning attacks can compromise the safety of large language models (LLMs) by injecting malicious documents into their training data. Existing work has studied pretraining poisoni…
cs.CL2024
Prompto: An open source library for asynchronous querying of LLM endpoints
Ryan Sze-Yin Chan, Federico Nanni, Angus R. Williams +7
Recent surge in Large Language Model (LLM) availability has opened exciting avenues for research. However, efficiently interacting with these models presents a significant hurdle s…