activity
20242026
most citedHumanity's Last Exam

18 citations · 18 across the 1 of their papers we have counts for

collaborators

11 papers

cs.LG202618 cited

Humanity's Last Exam

Long Phan, Alice Gatti, Ziwen Han +1144

Benchmarks are important tools for tracking the rapid advancements in large language model (LLM) capabilities. However, benchmarks are not keeping pace in difficulty: LLMs now achi…

cs.LG2025

Online Convex Optimisation: The Optimal Switching Regret for all Segmentations Simultaneously

Stephen Pasteris, Chris Hicks, Vasilios Mavroudis +1

We consider the classic problem of online convex optimisation. Whereas the notion of static regret is relevant for stationary problems, the notion of switching regret is more appro…

cs.LG2025

Entity-based Reinforcement Learning for Autonomous Cyber Defence

Isaac Symes Thompson, Alberto Caron, Chris Hicks +1

A significant challenge for autonomous cyber defence is ensuring a defensive agent's ability to generalise across diverse network topologies and configurations. This capability is…

cs.LG2024

Extraction Propagation

Stephen Pasteris, Chris Hicks, Vasilios Mavroudis

Running backpropagation end to end on large neural networks is fraught with difficulties like vanishing gradients and degradation. In this paper we present an alternative architect…

cs.CR2024

Zero-Trust Network Access (ZTNA)

Vasilios Mavroudis

Zero-Trust Network Access (ZTNA) marks a significant shift in network security by adopting a "never trust, always verify" approach. This work provides an in-depth analysis of ZTNA,…

cs.CR2024

CybORG++: An Enhanced Gym for the Development of Autonomous Cyber Agents

Harry Emerson, Liz Bates, Chris Hicks +1

CybORG++ is an advanced toolkit for reinforcement learning research focused on network defence. Building on the CAGE 2 CybORG environment, it introduces key improvements, including…