4 papers
GPT-5 at CTFs: Case Studies From Top-Tier Cybersecurity Events
Reworr, Artem Petrov, Dmitrii Volkov
OpenAI and DeepMind's AIs recently got gold at the IMO math olympiad and ICPC programming competition. We show frontier AI is similarly good at hacking by letting GPT-5 compete in…
Paired-Sampling Contrastive Framework for Joint Physical-Digital Face Attack Detection
Andrei Balykin, Anvar Ganiev, Denis Kondranin +3
Modern face recognition systems remain vulnerable to spoofing attempts, including both physical presentation attacks and digital forgeries. Traditionally, these two attack vectors…
Evaluating AI cyber capabilities with crowdsourced elicitation
Artem Petrov, Dmitrii Volkov
As AI systems become increasingly capable, understanding their offensive cyber potential is critical for informed governance and responsible deployment. However, it's hard to accur…
Hacking CTFs with Plain Agents
Rustem Turtayev, Artem Petrov, Dmitrii Volkov +1
We saturate a high-school-level hacking benchmark with plain LLM agent design. Concretely, we obtain 95% performance on InterCode-CTF, a popular offensive security benchmark, using…